EditHero: A Benchmark for Long-Horizon Part-Level 3D Editing and Vibe Modeling Paper • 2610.02298 • Published 8 days ago • 57
More Choices, Fewer Decisions: Ordinal-Scale Bias in JEV-like Direct-Decision Models Paper • 2609.38827 • Published 9 days ago • 60
EvoDuet: Bilevel Co-Evolution of Web Searching and Task Solving for Scientific Discovery Paper • 2609.40340 • Published 9 days ago • 110
HybridCUA: Learning to Orchestrate GUI and CLI for Computer-Use Agents Paper • 2609.38008 • Published 10 days ago • 47
Disaggregated Quantization: Specializing LLM Prefill and Decode Paper • 2609.26333 • Published 17 days ago • 93
OmniEcho: Spatial Audio Understanding for Embodied Agents Paper • 2609.23407 • Published 19 days ago • 35
Schrödinger's Code Repository: Have LLMs Learned SWE-bench or Memorized It? Paper • 2609.27891 • Published Aug 21 • 32
Confidence Comes from Experience: Experiential Confidence Estimation from Reasoning to Agents Paper • 2609.17708 • Published 24 days ago • 78
GAE: Learning a Geometry-Native Latent Space for 3D-Consistent World Generation Paper • 2609.24981 • Published 18 days ago • 76
From Pattern Recognizers to Personalized Companions: A Survey of Large Language Models in Mental Health Paper • 2609.25186 • Published 18 days ago • 32
Ovis-Embedding: Pushing the Frontiers of Universal Omni-Modal Embeddings Paper • 2609.25165 • Published 18 days ago • 77
Tri-PvP: Exposing Modality Bias in Omni-Modal Large Language Models through Perceptual-Propositional Evidence Conflicts Paper • 2609.06011 • Published Sep 5 • 18
ALPINE: Adaptive Localization for Parameter- and Sample-Efficient Few-Shot Learning Paper • 2609.22323 • Published 23 days ago • 10
RecreationWorld: Scalable and Verifiable Environments for Hybrid Computer-Use Agents Paper • 2609.22000 • Published 21 days ago • 80