open-llm-leaderboard-old/details_alexredna__Tukan-1.1B-Chat-reasoning-sft-COLA Updated Jan 22, 2024 • 273 • 5
AutoGUIWorld: Image Generators as Visual World Models for GUI Agent Paper • 2610.01215 • Published 8 days ago • 61
Surprising Success, Repeated Failure: Entropy-Guided Credit Assignment for Exploration in LLM Reasoning Paper • 2609.33781 • Published 12 days ago • 46
Raven: The Harness of Harnesses for Composable Agentic Intelligence Paper • 2609.33439 • Published 12 days ago • 569
CreitinGameplays/magpie-reasoning-v1-10k-step-by-step-rationale-alpaca-format-changedtoken Viewer • Updated Feb 9, 2025 • 10k • 84 • 6
APM-Bench: Benchmarking Cross-session Persistent Memory for Egocentric Streaming Video Assistants Paper • 2609.37559 • Published 10 days ago • 46
An RL View of OPD: Least Square Policy Distillation for Sample-Efficient LLM Reasoning Paper • 2609.35505 • Published 11 days ago • 26
tungvu3196/vlm-project-with-images-with-bbox-images-with-tree-of-thoughts-with-original Viewer • Updated Sep 30, 2025 • 12.6k • 83 • 4
DeL-TaiseiOzaki/Tengentoppa-llm-jp-13B-reasoning-it Text Generation • 14B • Updated Dec 15, 2024 • 146 • 8
FuseReg: Regularizing Layer Fusion Mitigates the Reconstruction-Generation Gap in Representation Autoencoders Paper • 2609.31620 • Published 14 days ago • 160
LangAGI-Lab/magpie-reasoning-v1-20k-math-verifiable-step-by-step-rationale-alpaca-format Viewer • Updated Feb 10, 2025 • 20k • 276 • 10
tungvu3196/vlm-project-with-images-with-bbox-images-with-tree-of-thoughts-v2 Viewer • Updated Jun 15, 2025 • 12.3k • 113 • 5
LangAGI-Lab/magpie-reasoning-v1-20k-math-verifiable-step-by-step-rationale Viewer • Updated Feb 10, 2025 • 20k • 180 • 10
WanPE: Towards Cinematic Prompt Enhancement for Modern Text-to-Video Generation Paper • 2609.30221 • Published 15 days ago • 47