Looping Beyond Twice: A Scalable Recipe for Looped Mixture-of-Experts Paper • 2610.01153 • Published 7 days ago • 22
Science Utopia? Closed-Loop LLM Simulation of Academic Research Ecosystems Paper • 2610.01257 • Published 7 days ago • 44
LexReward: A Taxonomy-Driven Reward Framework for Legal Language Models Paper • 2609.39071 • Published 8 days ago • 63
RealCompanion: Benchmarking Human Understanding from Reasoning over Longitudinal Real-World Conversations Paper • 2610.01780 • Published 7 days ago • 275
Pivot-SD: Efficient Self-Distillation for Masked Diffusion Language Models Paper • 2610.03665 • Published 6 days ago • 59
Raven: The Harness of Harnesses for Composable Agentic Intelligence Paper • 2609.33439 • Published 11 days ago • 569
Scaling Properties of Same-Family On-Policy Distillation Paper • 2609.32722 • Published 12 days ago • 325