WarpSAC: Towards the Pinnacle of Scalable Off-policy RL by Rethinking Exploration and Exploitation Paper • 2608.24479 • Published 3 days ago • 134
RecHarness: A Bandit-Routed Agentic Harness for Self-Evolving Recommender Systems Paper • 2607.29241 • Published 28 days ago • 11
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents Paper • 2607.28227 • Published 29 days ago • 309
Flow-ERD: Agent-type Aware Flow Matching with Entropy-Regularized Distillation for Diverse Traffic Simulation Paper • 2607.06957 • Published Jul 8 • 12
AnyBokeh: Physics-Guided Any-to-Any Bokeh Editing with Optical Fingerprint Transfer Paper • 2606.31959 • Published Jun 30 • 10
Geometric Stability of Neural Population Codes: Regional Variation, Behavioral Relevance, and Circuit Dependence Paper • 2606.29655 • Published Jun 28 • 4
LongDS-Bench: On the Failure of Long-Horizon Agentic Data Analysis Paper • 2605.30434 • Published May 28 • 23