yuxinlu1/gemma-4-12B-coder-fable5-composer2.5-v1-GGUF Text Generation • 12B • Updated Jun 19 • 291k • 2.82k
sakamakismile/gemma-4-12B-coder-fable5-composer2.5-MTP-NVFP4 Text Generation • 7B • Updated Jun 16 • 1.5k • 49
sakamakismile/gemma-4-12B-coder-fable5-composer2.5-GGUF Text Generation • 12B • Updated Jun 16 • 320 • 5
Kimuraxhalu/gemma-4-12B-coder-fable5-composer2.5-MTP-NVFP4 Text Generation • 7B • Updated Jun 16 • 55 • 3
HauhauCS/Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive Image-Text-to-Text • 35B • Updated Apr 17 • 2.01M • 3.38k
DavidAU/Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-NEO-CODE-Di-IMatrix-MAX-GGUF Image-Text-to-Text • 39B • Updated 22 days ago • 240k • 696
mlx-community/diffusiongemma-26B-A4B-it-4bit Image-Text-to-Text • 5B • Updated 27 days ago • 2.05k • 35
deadbydawn101/ravenx-Gemma4-12B-MTP-OBLITERATED-OpenMAI-OpenMythos-deep-reasoning-GGUF Text Generation • 12B • Updated Jun 11 • 337 • 15
stamsam/Qwen3.6-35B-A3B-Claude-4.7-Opus-Reasoning-Distilled-MLX-oQ4-MTP Text Generation • 35B • Updated May 12 • 3.57k • 25
srv-sngh/gemma-4-12B-coder-fable5-composer2.5-nvfp4 Text Generation • 2B • Updated Jun 25 • 1.05k • 10
Production-Grade Local LLM Inference on Apple Silicon: A Comparative Study of MLX, MLC-LLM, Ollama, llama.cpp, and PyTorch MPS Paper • 2511.05502 • Published Oct 9, 2025
In-the-Flow Agentic System Optimization for Effective Planning and Tool Use Paper • 2510.05592 • Published Oct 7, 2025 • 113