CodeMidas: Scaling Agentic Coding RL Environments from Code Itself Paper • 2609.22068 • Published 10 days ago • 136
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses Paper • 2609.24972 • Published 7 days ago • 211
Transferring the Intelligence of VLMs to Robotic Control Paper • 2609.22966 • Published 9 days ago • 119