Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
26.2
TFLOPS
ddh0
ddh0
93
107
1080
Follow
richardbenson91477's profile picture
cmh's profile picture
VikashSingh's profile picture
171 followers
·
103 following
ddh0
AI & ML interests
None yet
Recent Activity
reacted
to
eaddario
's
post
with ❤️
1 day ago
Experimental global target bits‑per‑weight quantization of **XHToken/Spark-X2.5-1.7B** and **XHToken/Spark-X2.5-4B**. Unlike standard llama.cpp quantization that rely on fixed type heuristics (e.g., Q4_K_M), the Target BPW approach automatically optimizes per-tensor precision where it matters the most, and produces high quality models that meet a precise global size target. Key Advantages: - VRAM Maximization: Can generate high quality models sized exactly to fit hardware constraints (e.g., fitting the model into exactly 24GB VRAM). - Data-Driven Precision: Quantization mix is determined by actual weight error sensitivity rather than hardcoded rules, often yielding better PPL/KLD size trade-offs. Full benchmarks (PPL, KLD, ARC, GPQA, MMLU, etc.) and methodology in the model's card. https://huggingface.co/eaddario/Spark-X2.5-1.7B-GGUF https://huggingface.co/eaddario/Spark-X2.5-4B-GGUF
reacted
to
eaddario
's
post
with 🔥
1 day ago
Experimental global target bits‑per‑weight quantization of **XHToken/Spark-X2.5-1.7B** and **XHToken/Spark-X2.5-4B**. Unlike standard llama.cpp quantization that rely on fixed type heuristics (e.g., Q4_K_M), the Target BPW approach automatically optimizes per-tensor precision where it matters the most, and produces high quality models that meet a precise global size target. Key Advantages: - VRAM Maximization: Can generate high quality models sized exactly to fit hardware constraints (e.g., fitting the model into exactly 24GB VRAM). - Data-Driven Precision: Quantization mix is determined by actual weight error sensitivity rather than hardcoded rules, often yielding better PPL/KLD size trade-offs. Full benchmarks (PPL, KLD, ARC, GPQA, MMLU, etc.) and methodology in the model's card. https://huggingface.co/eaddario/Spark-X2.5-1.7B-GGUF https://huggingface.co/eaddario/Spark-X2.5-4B-GGUF
updated
a model
1 day ago
ddh0/imatrices
View all activity
Organizations
ddh0
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
teto3/DeepSeek-V4-Flash-Base-Q4KExperts-GGUF
about 1 month ago
Hello, how did you convert this model to GGUF?
1
#1 opened about 1 month ago by
ddh0
New activity in
ddh0/DeepSeek-V4-Flash-GGUF
about 2 months ago
MTP with DeepSeek-V4-Flash-0731
👍
1
1
#1 opened about 2 months ago by
anikifoss
New activity in
ddh0/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16-GGUF
about 2 months ago
nemotron
#1 opened about 2 months ago by
willowoods
New activity in
Gryphe/Gemma-4-31B-StyleTune
3 months ago
This is a very cool idea!
👍
1
1
#3 opened 3 months ago by
ddh0
New activity in
google/gemma-4-12B-it
4 months ago
fix max_position_embeddings 128Ki --> 256Ki
1
#11 opened 4 months ago by
ddh0
New activity in
mistralai/Mistral-Small-4-119B-2603
6 months ago
Recommended sampler?
4
#4 opened 6 months ago by
mratsim
New activity in
zai-org/GLM-4.7-Flash
8 months ago
llama.cpp inference - 20 times (!) slower than OSS 20 on a RTX 5090
➕
1
9
#12 opened 8 months ago by
cmp-nct
New activity in
ddh0/dolphin-2.1-mistral-7b-GGUF-fp16
8 months ago
Update README.md
#1 opened 8 months ago by
cherry0328
New activity in
ddh0/Cassiopeia-70B
9 months ago
This Model Rules
❤️
1
1
#2 opened 9 months ago by
Firworks
New activity in
TheDrummer/Snowpiercer-15B-v4-GGUF
9 months ago
Shittiest Performing Model
15
#1 opened 9 months ago by
ialhabbal
New activity in
bartowski/Qwen_Qwen3-VL-30B-A3B-Instruct-GGUF
10 months ago
Problem
6
#1 opened 11 months ago by
NeKonnnn
New activity in
ilintar/Qwen3-Next-80B-A3B-Instruct-GGUF
11 months ago
Fix model name (not A30B, but A3B)
#1 opened 11 months ago by
ddh0
New activity in
zai-org/GLM-4.5-Air
11 months ago
Will GLM-4.6-Air model be released?
❤️
🚀
9
8
#15 opened 11 months ago by
ddh0
New activity in
QuantPasture/GLM-4.6-GGUF
12 months ago
What is the .bin file?
7
#1 opened 12 months ago by
Downtown-Case
New activity in
bartowski/zai-org_GLM-4.6-GGUF
12 months ago
Quants without imatrix
17
#2 opened 12 months ago by
Rotating
New activity in
MikeRoz/GLM-4.6-exl3
12 months ago
2.25 bpw perplexity
20
#2 opened 12 months ago by
malamen4
New activity in
QuantPasture/GLM-4.5-GGUF
12 months ago
clarify text corpus version
#1 opened 12 months ago by
ddh0
New activity in
jukofyork/creative-writing-control-vectors-v3.0
about 1 year ago
Wur doomed!
566
#14 opened over 1 year ago by
jukofyork
New activity in
TheDrummer/GLM-Steam-106B-A12B-v1-GGUF
about 1 year ago
I love it
❤️
2
#1 opened about 1 year ago by
ddh0
New activity in
ggml-org/Kimi-VL-A3B-Thinking-2506-GGUF
about 1 year ago
Unable to load mmproj: `load_hparams: unknown projector type: kimivl`
2
#1 opened about 1 year ago by
ddh0
Load more