PCSD: Persistent Consistency for Self-Distillation in Agentic Reinforcement Learning Paper • 2608.01837 • Published Aug 3 • 39
ATOD: Annealed Turn-aware On-policy Distillation for Multi-turn Autonomous Agents Paper • 2606.27814 • Published Jun 26 • 1
SPA-RL: Reinforcing LLM Agents via Stepwise Progress Attribution Paper • 2505.20732 • Published May 27, 2025 • 2
Dr. MAS: Stable Reinforcement Learning for Multi-Agent LLM Systems Paper • 2602.08847 • Published Feb 9 • 30
LLM-JEPA: Large Language Models Meet Joint Embedding Predictive Architectures Paper • 2509.14252 • Published Sep 11, 2025 • 10
VLA-JEPA: Enhancing Vision-Language-Action Model with Latent World Model Paper • 2602.10098 • Published Feb 10 • 24
A-JEPA: Joint-Embedding Predictive Architecture Can Listen Paper • 2311.15830 • Published Nov 27, 2023 • 1
ThinkJEPA: Empowering Latent World Models with Large Vision-Language Reasoning Model Paper • 2603.22281 • Published Mar 23 • 22
Causal-JEPA: Learning World Models through Object-Level Latent Interventions Paper • 2602.11389 • Published Feb 11 • 13
VL-JEPA: Joint Embedding Predictive Architecture for Vision-language Paper • 2512.10942 • Published Dec 11, 2025 • 63
OPD-V: Visual On-Policy Self-Distillation with Modality Balance Paper • 2608.05131 • Published Aug 6 • 14
V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning Paper • 2506.09985 • Published Jun 11, 2025 • 35
LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics Paper • 2511.08544 • Published Nov 11, 2025 • 14