YaRN: Efficient Context Window Extension of Large Language Models Paper • 2309.00071 • Published Aug 31, 2023 • 87
view article Article Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps +1 iamleonie, burtenshaw, sergiopaniego • 21 days ago • 132
LayerTracer: Cognitive-Aligned Layered SVG Synthesis via Diffusion Transformer Paper • 2502.01105 • Published Feb 3, 2025 • 22
H3-World: Turning Language Understanding into World Control Paper • 2609.01560 • Published 23 days ago • 52
GameEngineBench: Evaluating Coding Agents on Real C++ Runtime Environments Paper • 2607.03525 • Published Jul 15 • 1
GameXpert-Bench: How Far Are Coding Agents from Expert Game Development? Paper • 2608.21833 • Published Aug 22 • 17
Agentic Game Development as a Verifiable Trajectory Data Engine for Scaling World Models Paper • 2608.25518 • Published 29 days ago • 59
How Your Credentials Are Leaked by LLM Agent Skills: An Empirical Study Paper • 2604.03070 • Published Jun 19 • 2
MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published Aug 4 • 54
view article Article Meta is back with Muse Glimmer: local, agentic, multimodal, and open source +2 pcuenq, merve, burtenshaw, ariG23498 • Aug 10 • 113
Qwen-AgentWorld: Language World Models for General Agents Paper • 2606.24597 • Published Jun 23 • 165
Hackphyr: A Local Fine-Tuned LLM Agent for Network Security Environments Paper • 2409.11276 • Published Sep 17, 2024 • 11
Laguna S 2.1 Collection Our most capable model to date, designed for long-horizon work. • 13 items • Updated Aug 3 • 51
view article Article Training a 2.7B MoE from scratch for $200, one GPU at a time vovaRL • Jul 29 • 6