view article Article Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps +1 iamleonie, burtenshaw, sergiopaniego • 15 days ago • 121
view article Article BenchMIRT: What are LLM benchmarks actually measuring? allenai • 16 days ago • 22
view article Article Measuring benchmark optimization in speech recognition +5 tlebryk02, bezzam, aliceebaird, dayllon, jpc, jens-hume-ai, tzirakis • 28 days ago • 66
view article Article State of Open Models: Summer 2026 Observations +1 AdinaY, multimodalart, irenesolaiman • Aug 14 • 203
Running 6 Transformers Model Architectures 📐 6 Browse and filter transformer model architecture diagrams
view article Article Introducing North Mini Code: Cohere’s First Model For Developers CohereLabs • Jun 9 • 86
nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 Text Generation • 561B • Updated 24 days ago • 176k • • 345