Pre-selection of models to consider across languages for Alpha - Only the Cohere models are currently served by Inference Providers
Yacine Jernite
AI & ML interests
Technical, community, and regulatory tools of AI governance @HuggingFace
Recent Activity
updated a collection about 9 hours ago
Model picker - potluck liked a model about 9 hours ago
ssc-dsai/gc-llm-apertus-70b-instruct-2509 liked a Space about 14 hours ago
enjalot/latent-craft-blOrganizations
Text embedding models I use
-
google/embeddinggemma-300m
Sentence Similarity • 0.3B • Updated • 2.23M • • 1.89k -
Qwen/Qwen3-Embedding-0.6B
Feature Extraction • 0.6B • Updated • 7.5M • • 1.19k -
Qwen/Qwen3-Embedding-4B-GGUF
4B • Updated • 41.6k • 124 -
ibm-granite/granite-embedding-english-r2
Feature Extraction • 0.1B • Updated • 50.5k • 87
Models for local deployment
Document processing
Cybersecurity
super-smol to fine-tune
Inference-supported production models
List of recent models to use through HF inference providers
-
CohereLabs/command-a-vision-07-2025
Image-Text-to-Text • 112B • Updated • 18.2k • • 88 -
Qwen/Qwen3-235B-A22B-Instruct-2507
Text Generation • 235B • Updated • 147k • • 796 -
Qwen/Qwen3-32B
Text Generation • 33B • Updated • 5.16M • • 744 -
openai/gpt-oss-120b
Text Generation • 117B • Updated • 5.41M • • 5.18k
Privacy
Model picker - potluck
Pre-selection of models to consider across languages for Alpha - Only the Cohere models are currently served by Inference Providers
Cybersecurity
Text embedding models I use
-
google/embeddinggemma-300m
Sentence Similarity • 0.3B • Updated • 2.23M • • 1.89k -
Qwen/Qwen3-Embedding-0.6B
Feature Extraction • 0.6B • Updated • 7.5M • • 1.19k -
Qwen/Qwen3-Embedding-4B-GGUF
4B • Updated • 41.6k • 124 -
ibm-granite/granite-embedding-english-r2
Feature Extraction • 0.1B • Updated • 50.5k • 87
super-smol to fine-tune
Models for local deployment
Inference-supported production models
List of recent models to use through HF inference providers
-
CohereLabs/command-a-vision-07-2025
Image-Text-to-Text • 112B • Updated • 18.2k • • 88 -
Qwen/Qwen3-235B-A22B-Instruct-2507
Text Generation • 235B • Updated • 147k • • 796 -
Qwen/Qwen3-32B
Text Generation • 33B • Updated • 5.16M • • 744 -
openai/gpt-oss-120b
Text Generation • 117B • Updated • 5.41M • • 5.18k
Document processing
Privacy