Agora: Git as Shared Memory for Collective AutoResearch Paper • 2609.18094 • Published 9 days ago • 57
VC-Attention: Value Smoothing and Softmax Casting for Low-bit Attention Paper • 2609.15810 • Published 11 days ago • 50
FLAT: Resampling Image and Text into 1D Flexible-Length Aligned Transmodal Tokens for Retrieval and Generation Paper • 2609.16591 • Published 10 days ago • 16
Learning to Solve Hard Problems in RL for LLMs by Never Giving Up Paper • 2609.13443 • Published 14 days ago • 13
Story Imprinting: AI Assistants Absorb Traits from Human Characters They Resemble Paper • 2609.10883 • Published 16 days ago • 1
TurnBench: A Multi-Domain Benchmark for Turn-Taking Dynamics in Spoken Dialogue Paper • 2608.25218 • Published Aug 25 • 1
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 11 days ago • 246
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published 22 days ago • 186
Don't Drop Dropout: Optimizing Layer Sparsity for Efficient LLM Training and Inference Paper • 2609.05275 • Published 21 days ago • 26
Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation Paper • 2609.11638 • Published 15 days ago • 700
Jina-OCR-v1: Efficient Document Parsing with Speculative Decoding and Dense Verifiable Rewards Paper • 2609.03181 • Published 23 days ago • 11
Quantization-Aware Healing: A Practical Recipe for Recovering Compressed, 4-Bit LLMs Paper • 2608.20953 • Published Aug 21 • 14
Let's Scale Step by Step: Compute-Efficient Hyperparameter Transfer for Large-Scale Mixture-of-Experts Paper • 2608.20061 • Published Aug 20 • 47
SkillAdam: Stable and Efficient Skill Evolution for Agents Paper • 2609.08944 • Published 17 days ago • 12
PARSER: Read in Parallel, Reason in Depth for Long-Context LLM Agents Paper • 2609.06702 • Published 19 days ago • 26