Negative Self-Distillation: Learning to Reason by Avoiding Flaws Paper • 2609.11699 • Published 3 days ago • 17
Multi-Grid Post-Training for Long-Form Multi-Shot Video Generation Paper • 2609.06373 • Published 7 days ago • 21
Mask Forcing: Improving Autoregressive Video Diffusion Distillation via Dual-Noise Masking Rollout Paper • 2609.09123 • Published 5 days ago • 49
Locked at the Entrance, Open Inside: Where RLVR Narrows the Solution Space Paper • 2608.29188 • Published 15 days ago • 10
Knowing When Not to Reuse: Conditional Experience Transfer in Autonomous LLM Post-Training Paper • 2608.26730 • Published 17 days ago • 155
SemComp-Bench: Benchmarking Semantic Task Completion in Video Generation Paper • 2608.17426 • Published 26 days ago • 159
AutoDesign: Meta-Harness Optimization for Long-Horizon Agentic Design Paper • 2608.13560 • Published Aug 13 • 63
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 263
SkillJack: Persistent Skill Backdoors in Self-Evolving Agents Paper • 2608.03509 • Published Aug 4 • 24