Pre-selection of models to consider across languages for Alpha - Only the Cohere models are currently served by Inference Providers
Yacine Jernite
AI & ML interests
Technical, community, and regulatory tools of AI governance @HuggingFace
Recent Activity
updated a collection 5 days ago
Model picker - potluck liked a model 5 days ago
ssc-dsai/gc-llm-apertus-70b-instruct-2509 liked a Space 5 days ago
enjalot/latent-craft-blOrganizations
Text embedding models I use
-
google/embeddinggemma-300m
Sentence Similarity • 0.3B • Updated • 2.11M • • 1.91k -
Qwen/Qwen3-Embedding-0.6B
Feature Extraction • 0.6B • Updated • 8.35M • • 1.22k -
Qwen/Qwen3-Embedding-4B-GGUF
4B • Updated • 41.4k • 127 -
ibm-granite/granite-embedding-english-r2
Feature Extraction • 0.1B • Updated • 56.5k • 87
Models for local deployment
Document processing
Cybersecurity
super-smol to fine-tune
Inference-supported production models
List of recent models to use through HF inference providers
-
CohereLabs/command-a-vision-07-2025
Image-Text-to-Text • 112B • Updated • 17.7k • • 88 -
Qwen/Qwen3-235B-A22B-Instruct-2507
Text Generation • 235B • Updated • 166k • • 798 -
Qwen/Qwen3-32B
Text Generation • 33B • Updated • 4.86M • • 747 -
openai/gpt-oss-120b
Text Generation • 117B • Updated • 5.42M • • 5.23k
Privacy
Model picker - potluck
Pre-selection of models to consider across languages for Alpha - Only the Cohere models are currently served by Inference Providers
Cybersecurity
Text embedding models I use
-
google/embeddinggemma-300m
Sentence Similarity • 0.3B • Updated • 2.11M • • 1.91k -
Qwen/Qwen3-Embedding-0.6B
Feature Extraction • 0.6B • Updated • 8.35M • • 1.22k -
Qwen/Qwen3-Embedding-4B-GGUF
4B • Updated • 41.4k • 127 -
ibm-granite/granite-embedding-english-r2
Feature Extraction • 0.1B • Updated • 56.5k • 87
super-smol to fine-tune
Models for local deployment
Inference-supported production models
List of recent models to use through HF inference providers
-
CohereLabs/command-a-vision-07-2025
Image-Text-to-Text • 112B • Updated • 17.7k • • 88 -
Qwen/Qwen3-235B-A22B-Instruct-2507
Text Generation • 235B • Updated • 166k • • 798 -
Qwen/Qwen3-32B
Text Generation • 33B • Updated • 4.86M • • 747 -
openai/gpt-oss-120b
Text Generation • 117B • Updated • 5.42M • • 5.23k
Document processing
Privacy