Apple Silicon (MLX) builds of GLM-5.3 (744B) with a fixed glm_moe_dsa runtime: github.com/PipeNetwork/glm53-mlx
Pipenetwork PRO
pipenetwork
AI & ML interests
AI & ML
Recent Activity
updated a model 4 days ago
pipenetwork/GLM-5.2-REAP50-MLX-4bit updated a model 4 days ago
pipenetwork/GLM-5.2-REAP37-MLX-4bit updated a model 4 days ago
pipenetwork/GLM-5.2-REAP25-MLX-4bitOrganizations
None yet
Qwen3.8-Flash-Next MLX
Apple Silicon (MLX) builds of Qwen3.8-Flash-Next with a validated qwen4_exp runtime: github.com/PipeNetwork/qwen38-flash-next-mlx
-
pipenetwork/Qwen3.8-Flash-Next-MLX-8bit
Text Generation • 52B • Updated • 857 -
pipenetwork/Qwen3.8-Flash-Next-MLX-6bit
Text Generation • 41B • Updated • 1.15k • 1 -
pipenetwork/Qwen3.8-Flash-Next-MLX-mixed-4_8bit
Text Generation • 30B • Updated • 3.65k • 7 -
pipenetwork/Qwen3.8-Flash-Next-MLX-4bit
Text Generation • 30B • Updated • 893
Muse-Glimmer 30B MLX
Apple Silicon (MLX) build of Muse-Glimmer-30B, plus the runtime no released mlx-vlm has: github.com/PipeNetwork/muse-glimmer-mlx
Inkling-Small · MLX
276B-A12B multimodal MoE on Apple Silicon. Same code path as Inkling 975B. REAP-25 is free: github.com/PipeNetwork/inkling-mlx
-
pipenetwork/Inkling-Small-MLX-8bit
Image-Text-to-Text • 74B • Updated • 104 -
pipenetwork/Inkling-Small-MLX-6bit
Image-Text-to-Text • 58B • Updated • 82 -
pipenetwork/Inkling-Small-MLX-4bit
Image-Text-to-Text • 41B • Updated • 321 • 2 -
pipenetwork/Inkling-Small-MLX-REAP25-4bit
Image-Text-to-Text • 31B • Updated • 295 • 2
DeepSeek-V4-Flash · MLX
304B MoE on Apple Silicon. deepseek_v4 is in no released runtime — ported from scratch. Take mixed-4_8bit. github.com/PipeNetwork/deepseek-v4-mlx
-
pipenetwork/DeepSeek-V4-Flash-MLX-8bit
Text Generation • 81B • Updated • 361 -
pipenetwork/DeepSeek-V4-Flash-MLX-6bit
Text Generation • 64B • Updated • 213 -
pipenetwork/DeepSeek-V4-Flash-MLX-mixed-4_8bit
Text Generation • 47B • Updated • 409 • 1 -
pipenetwork/DeepSeek-V4-Flash-MLX-4bit
Text Generation • 46B • Updated • 312
Qwen3.6-35B-A3B — MLX
MLX (Apple Silicon) conversion of Qwen/Qwen3.6-35B-A3B in NVFP4 — the MLX analog of nvidia/Qwen3.6-35B-A3B-NVFP4. Text-only.
GLM-5.2 REAP (expert-pruned)
REAP expert-pruned GLM-5.2 (smaller/faster). REAP-25 near-lossless (+2.3% PPL); REAP-37 +7.3%; REAP-50 smallest but +37.5%.
VISTA MLX
First MLX (Apple Silicon) quantizations of inclusionAI VISTA-9B and VISTA-4B (qwen3_5).
Gemma-4-26B-A4B-it MLX
MLX quantizations of google/gemma-4-26B-A4B-it (MoE): 5/6/8-bit (4-bit available from mlx-community). Text-only.
Kimi-K2.7-Code MLX
MLX build of Kimi-K2.7-Code. Base is natively 4-bit (int4 experts + bf16 rest); this keeps experts at 4-bit and lifts non-expert layers to 6-bit.
Holo-3.1 MLX (computer-use)
First working MLX builds of H Company's Holo-3.1 vision-language computer-use agents (Qwen3.5-VL). Vision-validated. Apache-2.0.
-
pipenetwork/Holo-3.1-4B-MLX-4bit
Image-Text-to-Text • 1.0B • Updated • 81 • 3 -
pipenetwork/Holo-3.1-4B-MLX-8bit
Image-Text-to-Text • 2B • Updated • 19 • 1 -
pipenetwork/Holo-3.1-9B-MLX-4bit
Image-Text-to-Text • 2B • Updated • 24 -
pipenetwork/Holo-3.1-9B-MLX-8bit
Image-Text-to-Text • 3B • Updated • 36 • 1
Nemotron-3 MLX (Apple Silicon)
MLX quants of NVIDIA Nemotron-3 for Apple Silicon: Ultra 550B (4/5/6/8-bit) and dense Nano-4B (4/8-bit), converted with mlx-lm.
-
pipenetwork/NVIDIA-Nemotron-3-Ultra-550B-A55B-MLX-4bit
Text Generation • 549B • Updated • 121 -
pipenetwork/NVIDIA-Nemotron-3-Ultra-550B-A55B-MLX-8bit
Text Generation • 549B • Updated • 95 -
pipenetwork/NVIDIA-Nemotron-3-Nano-4B-MLX-8bit
1B • Updated • 2 -
pipenetwork/NVIDIA-Nemotron-3-Nano-4B-MLX-4bit
0.6B • Updated • 2
GLM-5.3-Flash MLX
Apple Silicon (MLX) builds of GLM-5.3-Flash with a validated, fixed glm5_next runtime: github.com/PipeNetwork/glm53-flash-mlx
-
pipenetwork/GLM-5.3-Flash-MLX-8bit
Image-Text-to-Text • 89B • Updated • 2.91k • 4 -
pipenetwork/GLM-5.3-Flash-MLX-6bit
Image-Text-to-Text • 69B • Updated • 1.82k • 1 -
pipenetwork/GLM-5.3-Flash-MLX-mixed-4_8bit
Image-Text-to-Text • 51B • Updated • 2.46k • 4 -
pipenetwork/GLM-5.3-Flash-MLX-4bit
Image-Text-to-Text • 50B • Updated • 3.06k • 2
Qwen3.8-2.4T-A95B MLX (REAP)
2.4T-param MoE for Apple Silicon. REAP-pruned because no standard quant fits 550GB. Every build measured per-layer against bf16.
MiniMax-H3 MLX
Apple Silicon (MLX) builds of MiniMax-H3, the 33B joint video+audio diffusion transformer. Code: github.com/PipeNetwork/minimax-h3-mlx
-
pipenetwork/MiniMax-H3-MLX-8bit
Image-Text-to-Video • 9B • Updated • 3.43k • 6 -
pipenetwork/MiniMax-H3-MLX-6bit
Image-Text-to-Video • 8B • Updated • 1.13k • 2 -
pipenetwork/MiniMax-H3-MLX-4bit
Image-Text-to-Video • 7B • Updated • 2.61k • 4 -
pipenetwork/MiniMax-H3-MLX-bf16
Image-Text-to-Video • 33B • Updated • 1.82k • 6
Inkling · MLX
975B-A41B multimodal MoE on Apple Silicon. Text+image+audio. Ported from scratch: github.com/PipeNetwork/inkling-mlx
Nemotron TwoTower · MLX
NVIDIA Nemotron TwoTower 30B-A3B diffusion LM in MLX for Apple Silicon: AR tower + two-tower diffusion, 4/6/8-bit + bf16.
-
pipenetwork/Nemotron-3-Nano-30B-A3B-context-mlx-4bit
Text Generation • 32B • Updated • 34 -
pipenetwork/Nemotron-3-Nano-30B-A3B-context-mlx-6bit
Text Generation • 32B • Updated • 41 • 1 -
pipenetwork/Nemotron-3-Nano-30B-A3B-context-mlx-8bit
Text Generation • 32B • Updated • 29 -
pipenetwork/Nemotron-3-Nano-30B-A3B-context-mlx-bf16
Text Generation • 32B • Updated • 40
Ornith-1.0-397B — MLX
MLX (Apple Silicon) conversions of deepreinforce-ai/Ornith-1.0-397B (Qwen3.5-MoE, text-only). 4/6/8-bit and bf16.
GLM-5.2 MLX
First MLX builds of zai-org/GLM-5.2 (glm_moe_dsa MoE): 4/5/6/8-bit + a 512GB-friendly mixed.
Rio-3.1-Open-30B MLX
First MLX (Apple Silicon) quantizations of prefeitura-rio/Rio-3.1-Open-30B (Qwen3-MoE): 4/5/6/8-bit.
Gemma-4-31B-it MLX
MLX (Apple Silicon) quantizations of google/gemma-4-31B-it: 4/5/6/8-bit. Text-only.
MiniMax-M3 MLX
MLX (Apple Silicon) text-only conversions of MiniMax-M3 (427B MoE): 3-bit to 8-bit plus a mixed-precision build.
-
pipenetwork/MiniMax-M3-MLX-8bit
Text Generation • 426B • Updated • 185 • 1 -
pipenetwork/MiniMax-M3-MLX-6bit
Text Generation • 426B • Updated • 232 • 1 -
pipenetwork/MiniMax-M3-MLX-4bit
Text Generation • 426B • Updated • 183 -
pipenetwork/MiniMax-M3-MLX-mixed-3_6bit
Text Generation • 426B • Updated • 242 • 2
Frog (SWE/debugging) MLX
MLX quants of Microsoft's FrogBoss-32B & FrogMini-14B (Qwen3 debugging finetunes, SWE-bench ~45% pass@1) for Apple Silicon.
-
pipenetwork/FrogMini-14B-2510-MLX-4bit
Text Generation • 15B • Updated • 12 -
pipenetwork/FrogMini-14B-2510-MLX-8bit
Text Generation • 15B • Updated • 6 -
pipenetwork/FrogBoss-32B-2510-MLX-4bit
Text Generation • 33B • Updated • 7 -
pipenetwork/FrogBoss-32B-2510-MLX-8bit
Text Generation • 33B • Updated • 7
GLM-5.3 MLX
Apple Silicon (MLX) builds of GLM-5.3 (744B) with a fixed glm_moe_dsa runtime: github.com/PipeNetwork/glm53-mlx
GLM-5.3-Flash MLX
Apple Silicon (MLX) builds of GLM-5.3-Flash with a validated, fixed glm5_next runtime: github.com/PipeNetwork/glm53-flash-mlx
-
pipenetwork/GLM-5.3-Flash-MLX-8bit
Image-Text-to-Text • 89B • Updated • 2.91k • 4 -
pipenetwork/GLM-5.3-Flash-MLX-6bit
Image-Text-to-Text • 69B • Updated • 1.82k • 1 -
pipenetwork/GLM-5.3-Flash-MLX-mixed-4_8bit
Image-Text-to-Text • 51B • Updated • 2.46k • 4 -
pipenetwork/GLM-5.3-Flash-MLX-4bit
Image-Text-to-Text • 50B • Updated • 3.06k • 2
Qwen3.8-Flash-Next MLX
Apple Silicon (MLX) builds of Qwen3.8-Flash-Next with a validated qwen4_exp runtime: github.com/PipeNetwork/qwen38-flash-next-mlx
-
pipenetwork/Qwen3.8-Flash-Next-MLX-8bit
Text Generation • 52B • Updated • 857 -
pipenetwork/Qwen3.8-Flash-Next-MLX-6bit
Text Generation • 41B • Updated • 1.15k • 1 -
pipenetwork/Qwen3.8-Flash-Next-MLX-mixed-4_8bit
Text Generation • 30B • Updated • 3.65k • 7 -
pipenetwork/Qwen3.8-Flash-Next-MLX-4bit
Text Generation • 30B • Updated • 893
Qwen3.8-2.4T-A95B MLX (REAP)
2.4T-param MoE for Apple Silicon. REAP-pruned because no standard quant fits 550GB. Every build measured per-layer against bf16.
Muse-Glimmer 30B MLX
Apple Silicon (MLX) build of Muse-Glimmer-30B, plus the runtime no released mlx-vlm has: github.com/PipeNetwork/muse-glimmer-mlx
MiniMax-H3 MLX
Apple Silicon (MLX) builds of MiniMax-H3, the 33B joint video+audio diffusion transformer. Code: github.com/PipeNetwork/minimax-h3-mlx
-
pipenetwork/MiniMax-H3-MLX-8bit
Image-Text-to-Video • 9B • Updated • 3.43k • 6 -
pipenetwork/MiniMax-H3-MLX-6bit
Image-Text-to-Video • 8B • Updated • 1.13k • 2 -
pipenetwork/MiniMax-H3-MLX-4bit
Image-Text-to-Video • 7B • Updated • 2.61k • 4 -
pipenetwork/MiniMax-H3-MLX-bf16
Image-Text-to-Video • 33B • Updated • 1.82k • 6
Inkling-Small · MLX
276B-A12B multimodal MoE on Apple Silicon. Same code path as Inkling 975B. REAP-25 is free: github.com/PipeNetwork/inkling-mlx
-
pipenetwork/Inkling-Small-MLX-8bit
Image-Text-to-Text • 74B • Updated • 104 -
pipenetwork/Inkling-Small-MLX-6bit
Image-Text-to-Text • 58B • Updated • 82 -
pipenetwork/Inkling-Small-MLX-4bit
Image-Text-to-Text • 41B • Updated • 321 • 2 -
pipenetwork/Inkling-Small-MLX-REAP25-4bit
Image-Text-to-Text • 31B • Updated • 295 • 2
Inkling · MLX
975B-A41B multimodal MoE on Apple Silicon. Text+image+audio. Ported from scratch: github.com/PipeNetwork/inkling-mlx
DeepSeek-V4-Flash · MLX
304B MoE on Apple Silicon. deepseek_v4 is in no released runtime — ported from scratch. Take mixed-4_8bit. github.com/PipeNetwork/deepseek-v4-mlx
-
pipenetwork/DeepSeek-V4-Flash-MLX-8bit
Text Generation • 81B • Updated • 361 -
pipenetwork/DeepSeek-V4-Flash-MLX-6bit
Text Generation • 64B • Updated • 213 -
pipenetwork/DeepSeek-V4-Flash-MLX-mixed-4_8bit
Text Generation • 47B • Updated • 409 • 1 -
pipenetwork/DeepSeek-V4-Flash-MLX-4bit
Text Generation • 46B • Updated • 312
Nemotron TwoTower · MLX
NVIDIA Nemotron TwoTower 30B-A3B diffusion LM in MLX for Apple Silicon: AR tower + two-tower diffusion, 4/6/8-bit + bf16.
-
pipenetwork/Nemotron-3-Nano-30B-A3B-context-mlx-4bit
Text Generation • 32B • Updated • 34 -
pipenetwork/Nemotron-3-Nano-30B-A3B-context-mlx-6bit
Text Generation • 32B • Updated • 41 • 1 -
pipenetwork/Nemotron-3-Nano-30B-A3B-context-mlx-8bit
Text Generation • 32B • Updated • 29 -
pipenetwork/Nemotron-3-Nano-30B-A3B-context-mlx-bf16
Text Generation • 32B • Updated • 40
Qwen3.6-35B-A3B — MLX
MLX (Apple Silicon) conversion of Qwen/Qwen3.6-35B-A3B in NVFP4 — the MLX analog of nvidia/Qwen3.6-35B-A3B-NVFP4. Text-only.
Ornith-1.0-397B — MLX
MLX (Apple Silicon) conversions of deepreinforce-ai/Ornith-1.0-397B (Qwen3.5-MoE, text-only). 4/6/8-bit and bf16.
GLM-5.2 REAP (expert-pruned)
REAP expert-pruned GLM-5.2 (smaller/faster). REAP-25 near-lossless (+2.3% PPL); REAP-37 +7.3%; REAP-50 smallest but +37.5%.
GLM-5.2 MLX
First MLX builds of zai-org/GLM-5.2 (glm_moe_dsa MoE): 4/5/6/8-bit + a 512GB-friendly mixed.
VISTA MLX
First MLX (Apple Silicon) quantizations of inclusionAI VISTA-9B and VISTA-4B (qwen3_5).
Rio-3.1-Open-30B MLX
First MLX (Apple Silicon) quantizations of prefeitura-rio/Rio-3.1-Open-30B (Qwen3-MoE): 4/5/6/8-bit.
Gemma-4-26B-A4B-it MLX
MLX quantizations of google/gemma-4-26B-A4B-it (MoE): 5/6/8-bit (4-bit available from mlx-community). Text-only.
Gemma-4-31B-it MLX
MLX (Apple Silicon) quantizations of google/gemma-4-31B-it: 4/5/6/8-bit. Text-only.
Kimi-K2.7-Code MLX
MLX build of Kimi-K2.7-Code. Base is natively 4-bit (int4 experts + bf16 rest); this keeps experts at 4-bit and lifts non-expert layers to 6-bit.
MiniMax-M3 MLX
MLX (Apple Silicon) text-only conversions of MiniMax-M3 (427B MoE): 3-bit to 8-bit plus a mixed-precision build.
-
pipenetwork/MiniMax-M3-MLX-8bit
Text Generation • 426B • Updated • 185 • 1 -
pipenetwork/MiniMax-M3-MLX-6bit
Text Generation • 426B • Updated • 232 • 1 -
pipenetwork/MiniMax-M3-MLX-4bit
Text Generation • 426B • Updated • 183 -
pipenetwork/MiniMax-M3-MLX-mixed-3_6bit
Text Generation • 426B • Updated • 242 • 2
Holo-3.1 MLX (computer-use)
First working MLX builds of H Company's Holo-3.1 vision-language computer-use agents (Qwen3.5-VL). Vision-validated. Apache-2.0.
-
pipenetwork/Holo-3.1-4B-MLX-4bit
Image-Text-to-Text • 1.0B • Updated • 81 • 3 -
pipenetwork/Holo-3.1-4B-MLX-8bit
Image-Text-to-Text • 2B • Updated • 19 • 1 -
pipenetwork/Holo-3.1-9B-MLX-4bit
Image-Text-to-Text • 2B • Updated • 24 -
pipenetwork/Holo-3.1-9B-MLX-8bit
Image-Text-to-Text • 3B • Updated • 36 • 1
Frog (SWE/debugging) MLX
MLX quants of Microsoft's FrogBoss-32B & FrogMini-14B (Qwen3 debugging finetunes, SWE-bench ~45% pass@1) for Apple Silicon.
-
pipenetwork/FrogMini-14B-2510-MLX-4bit
Text Generation • 15B • Updated • 12 -
pipenetwork/FrogMini-14B-2510-MLX-8bit
Text Generation • 15B • Updated • 6 -
pipenetwork/FrogBoss-32B-2510-MLX-4bit
Text Generation • 33B • Updated • 7 -
pipenetwork/FrogBoss-32B-2510-MLX-8bit
Text Generation • 33B • Updated • 7
Nemotron-3 MLX (Apple Silicon)
MLX quants of NVIDIA Nemotron-3 for Apple Silicon: Ultra 550B (4/5/6/8-bit) and dense Nano-4B (4/8-bit), converted with mlx-lm.
-
pipenetwork/NVIDIA-Nemotron-3-Ultra-550B-A55B-MLX-4bit
Text Generation • 549B • Updated • 121 -
pipenetwork/NVIDIA-Nemotron-3-Ultra-550B-A55B-MLX-8bit
Text Generation • 549B • Updated • 95 -
pipenetwork/NVIDIA-Nemotron-3-Nano-4B-MLX-8bit
1B • Updated • 2 -
pipenetwork/NVIDIA-Nemotron-3-Nano-4B-MLX-4bit
0.6B • Updated • 2