Self-OPD: On-Policy Distillation for Flow Matching Models without Teacher Paper • 2608.26872 • Published 9 days ago • 82
Agentic Game Development as a Verifiable Trajectory Data Engine for Scaling World Models Paper • 2608.25518 • Published 10 days ago • 196
Learn What's Left, Not What's Mastered: Saturation Aware Advantage Reweighting for Multi-Reward Policy Optimization Paper • 2608.16072 • Published 19 days ago • 151
Spark-to-Paper: End-to-End Research Paper Generation as a Composable Skill Paper • 2608.11924 • Published 24 days ago • 290
OneEmo: A Unified Multimodal Reasoning Model for Emotion Perception, Understanding, and Interaction Paper • 2608.06013 • Published about 1 month ago • 5
SimWAM: A Simple World Action Model for End-to-End Autonomous Driving Paper • 2608.07468 • Published 29 days ago • 107
Progress Reward Modeling for Robotic Learning: A Comprehensive Survey Paper • 2607.21655 • Published Jul 22 • 193
HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Paper • 2607.25895 • Published Jul 28 • 158
SUFLECA: Scaling Up Feature Learning for CAD-to-image Alignment Paper • 2607.15058 • Published Jul 16 • 8