Pre-selection of models to consider across languages for Alpha - Only the Cohere models are currently served by Inference Providers
Yacine Jernite
P(doom) Rejects the premise
AI & ML interests
Technical, community, and regulatory tools of AI governance @HuggingFace
Recent Activity
updated a collection about 16 hours ago
Model picker - potluck liked a model about 16 hours ago
Aleph-Alpha/Kolibri-1 upvoted an article about 19 hours ago
What a million datasets on Hugging Face tell us about open AI dataOrganizations
Text embedding models I use
-
google/embeddinggemma-300m
Sentence Similarity • 0.3B • Updated • 3.61M • • 1.98k -
Qwen/Qwen3-Embedding-0.6B
Feature Extraction • 0.6B • Updated • 9.25M • • 1.26k -
Qwen/Qwen3-Embedding-4B-GGUF
4B • Updated • 44.7k • 132 -
ibm-granite/granite-embedding-english-r2
Feature Extraction • 0.1B • Updated • 59k • 90
Models for local deployment
Document processing
Cybersecurity
super-smol to fine-tune
Inference-supported production models
List of recent models to use through HF inference providers
-
CohereLabs/command-a-vision-07-2025
Image-Text-to-Text • 112B • Updated • 19k • • 88 -
Qwen/Qwen3-235B-A22B-Instruct-2507
Text Generation • 235B • Updated • 178k • • 805 -
Qwen/Qwen3-32B
Text Generation • 33B • Updated • 3.7M • • 754 -
openai/gpt-oss-120b
Text Generation • 117B • Updated • 4.36M • • 5.36k
Privacy
Model picker - potluck
Pre-selection of models to consider across languages for Alpha - Only the Cohere models are currently served by Inference Providers
Cybersecurity
Text embedding models I use
-
google/embeddinggemma-300m
Sentence Similarity • 0.3B • Updated • 3.61M • • 1.98k -
Qwen/Qwen3-Embedding-0.6B
Feature Extraction • 0.6B • Updated • 9.25M • • 1.26k -
Qwen/Qwen3-Embedding-4B-GGUF
4B • Updated • 44.7k • 132 -
ibm-granite/granite-embedding-english-r2
Feature Extraction • 0.1B • Updated • 59k • 90
super-smol to fine-tune
Models for local deployment
Inference-supported production models
List of recent models to use through HF inference providers
-
CohereLabs/command-a-vision-07-2025
Image-Text-to-Text • 112B • Updated • 19k • • 88 -
Qwen/Qwen3-235B-A22B-Instruct-2507
Text Generation • 235B • Updated • 178k • • 805 -
Qwen/Qwen3-32B
Text Generation • 33B • Updated • 3.7M • • 754 -
openai/gpt-oss-120b
Text Generation • 117B • Updated • 4.36M • • 5.36k
Document processing
Privacy