TERRA: Terrain-Aware Reconstruction, Retargeting and Control for Musculoskeletal Locomotion Paper • 2609.38653 • Published 5 days ago • 33
Running 252 The ultimate guide to RL environments: building and scaling them in the LLM era 📝 252 Building and scaling RL environments for LLM training
No Prompt Left Behind: Exploiting Zero-Variance Prompts in LLM Reinforcement Learning via Entropy-Guided Advantage Shaping Paper • 2509.21880 • Published Sep 26, 2025 • 54