Pre-selection of models to consider across languages for Alpha - Only the Cohere models are currently served by Inference Providers
Yacine Jernite
AI & ML interests
Technical, community, and regulatory tools of AI governance @HuggingFace
Recent Activity
liked a model about 22 hours ago
Tele-AI/T1-115B updated a collection about 22 hours ago
Visualizations of AI Ecosystem Data liked a Space about 22 hours ago
AdinaY/china-open-models-download-trackerOrganizations
Text embedding models I use
-
google/embeddinggemma-300m
Sentence Similarity • 0.3B • Updated • 2.87M • • 1.94k -
Qwen/Qwen3-Embedding-0.6B
Feature Extraction • 0.6B • Updated • 9.14M • • 1.24k -
Qwen/Qwen3-Embedding-4B-GGUF
4B • Updated • 37k • 127 -
ibm-granite/granite-embedding-english-r2
Feature Extraction • 0.1B • Updated • 61.3k • 89
Models for local deployment
Document processing
Cybersecurity
super-smol to fine-tune
Inference-supported production models
List of recent models to use through HF inference providers
-
CohereLabs/command-a-vision-07-2025
Image-Text-to-Text • 112B • Updated • 19.1k • • 88 -
Qwen/Qwen3-235B-A22B-Instruct-2507
Text Generation • 235B • Updated • 211k • • 802 -
Qwen/Qwen3-32B
Text Generation • 33B • Updated • 5.06M • • 751 -
openai/gpt-oss-120b
Text Generation • 117B • Updated • 4.77M • • 5.3k
Privacy
Model picker - potluck
Pre-selection of models to consider across languages for Alpha - Only the Cohere models are currently served by Inference Providers
Cybersecurity
Text embedding models I use
-
google/embeddinggemma-300m
Sentence Similarity • 0.3B • Updated • 2.87M • • 1.94k -
Qwen/Qwen3-Embedding-0.6B
Feature Extraction • 0.6B • Updated • 9.14M • • 1.24k -
Qwen/Qwen3-Embedding-4B-GGUF
4B • Updated • 37k • 127 -
ibm-granite/granite-embedding-english-r2
Feature Extraction • 0.1B • Updated • 61.3k • 89
super-smol to fine-tune
Models for local deployment
Inference-supported production models
List of recent models to use through HF inference providers
-
CohereLabs/command-a-vision-07-2025
Image-Text-to-Text • 112B • Updated • 19.1k • • 88 -
Qwen/Qwen3-235B-A22B-Instruct-2507
Text Generation • 235B • Updated • 211k • • 802 -
Qwen/Qwen3-32B
Text Generation • 33B • Updated • 5.06M • • 751 -
openai/gpt-oss-120b
Text Generation • 117B • Updated • 4.77M • • 5.3k
Document processing
Privacy