陳能弘
Gavin-chen
AI & ML interests
None yet
Recent Activity
new activity about 4 hours ago
openbmb/MiniCPM5-1B-GGUF:Context is 4096, not the 131k advertised... new activity 1 day ago
deepgrove/maple-preview-GGUF:where are the benchmarks? liked a model 8 days ago
deepseek-ai/DeepSeek-V4.1-FlashOrganizations
None yet
Context is 4096, not the 131k advertised...
2
#9 opened 8 days ago
by
liar666
where are the benchmarks?
1
#9 opened 4 days ago
by
JNK333
Bro....500多B的Flash,8卡H200已经上不了桌了吗[cry]
9
#16 opened 8 days ago
by
Saito-Karuha
Can you make a 3:1 SWA version or linear attention version?
🔥 1
2
#6 opened 10 days ago
by
Gavin-chen
May I as
3
#1 opened 16 days ago
by
wh-wh-wh
GLM 5.3 Flash Lite idea
15
#13 opened 23 days ago
by
ChessVania
Will any other model get UD3 update?
➕👀 3
1
#108 opened 25 days ago
by
Gavin-chen
can you make gguf?
6
#1 opened about 1 month ago
by
Gavin-chen
Why Qwen3.8-27B overthinks? Here the reason and partial fix confirmed by benchmarks.
❤️👀 25
29
#76 opened about 1 month ago
by
LuffyTheFox
Why UD-Q2_K_XL has IQ1_M SSM alpha/beta and only token_embd use 2bit quant? It's a Q2 GGUF!
➕ 2
#36 opened about 1 month ago
by
Gavin-chen
I just checked Qwen3.8-2.4T-A95B for singular matrices and exploded condition numbers
👍 2
6
#29 opened about 1 month ago
by
LuffyTheFox
can you do 99999999999999B-A0.00000000000001B
🚀 2
5
#31 opened about 1 month ago
by
inikishev
Can't use bnb 4bit to load your model.
#11 opened about 2 months ago
by
Gavin-chen
Why so many empty response?
1
#15 opened 2 months ago
by
Gavin-chen
Nice Model, can beat Gemini3.1pro,sonnet5 , and GPT-5.5 instant
👍 2
1
#32 opened 3 months ago
by
Gavin-chen
how to use mtp in gemma gguf models??
4
#4 opened 3 months ago
by
koyukira
Any updates on the SWE test results?
👀 1
#10 opened 4 months ago
by
Gavin-chen
照顾一下Tesla T4的老人,需要8B、9B、14B、16B参数
2
#52 opened 5 months ago
by
zhousp666
I JUST LAUNCHED THIS AND IT BLOWN UP MY COMPUTER!!!!!!
2
#3 opened 5 months ago
by
uniquealexx
Can't wait anymore!
👍 2
2
#2 opened 5 months ago
by
Gavin-chen