Gerald Stanje
Gerald001
·
AI & ML interests
None yet
Recent Activity
new activity about 1 month ago
openai/gpt-oss-safeguard-20b:release the BF16 weights or nvfp4 new activity about 1 month ago
openai/gpt-oss-20b:MXFP4 utilization over NVFP4 new activity about 1 month ago
Qwen/Qwen3-TTS-12Hz-1.7B-CustomVoice:Inference latencyOrganizations
release the BF16 weights or nvfp4
#7 opened about 1 month ago
by
Gerald001
MXFP4 utilization over NVFP4
👍 2
2
#121 opened about 1 year ago
by
pprovins
Inference latency
1
#24 opened 7 months ago
by
popkek00
Question about inference speed
2
#10 opened 12 months ago
by
cX1y
I want transcription not translation
👍 5
2
#29 opened 8 months ago
by
smartire
What's the best way to run this in production for online serving?
2
#45 opened about 2 months ago
by
fikrikarim
Any plans for gpt oss 20b?
2
#1 opened 10 months ago
by
andhakanoon
Eagle3 for 20b
🚀 1
3
#4 opened 10 months ago
by
cr-boostrun
how to disable the reasoning mode?
👍 13
11
#50 opened about 1 year ago
by
szzzzz
GGUF is very slow for some reason
2
#12 opened 12 months ago
by
ineersa
Model conversion info
11
#9 opened 6 months ago
by
Gerald001
NVIDIA L40S GPU's for MXFP4 quantization
6
#100 opened about 1 year ago
by
lordim
How to turn off thinking mode
👍🔥 7
15
#86 opened about 1 year ago
by
Gierry
REASONING SETTING GUIDE 📚
😔👍 4
28
#28 opened about 1 year ago
by
xbruce22
Update chat_template.jinja
3
#229 opened 6 months ago
by
mohsin17444
question: setting reasoning effort
6
#66 opened about 1 year ago
by
TheBigBlockPC
Unable to load gpt-oss-20b on dual L40 (48GB) GPUs with vLLM
12
#136 opened 12 months ago
by
yjban
Model conversion info
11
#9 opened 6 months ago
by
Gerald001
assistantfinal, analysis keyword is contained in the huggingface gpt-oss-120 output. Is this intended?
👍 4
4
#130 opened 12 months ago
by
ml345