1,011 models · refreshed nightly
All models
Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.
| # | Model | Params | Context | Commercial use | 30d | Min VRAM |
|---|---|---|---|---|---|---|
| 151 | gemma-3-270m | 270 M | — | ⚠ gemma | 2.1 M | from 1.1 GB |
| 152 | faster-whisper-small | — | — | ✓ mit | 2.1 M | — |
| 153 | DeepSeek-OCR-2 | 3.4 B | 8 K | ✓ apache-2.0 | 2.1 M | from 8.5 GB |
| 154 | GLM-4.7-Flash | 31.2 B | 203 K | ✓ mit | 2.1 M | from 73.9 GB |
| 155 | embeddinggemma-300m | 300 M | — | ⚠ gemma | 2.1 M | from 1.9 GB |
| 156 | Meta-Llama-3-8B | 8.0 B | — | ⚠ llama3 | 2.1 M | from 19.4 GB |
| 157 | nemotron-3.5-asr-streaming-0.6b-gguf | — | — | other | 2.0 M | from 1.0 GB |
| 158 | diffusiongemma-26B-A4B-it | 25.8 B | — | ✓ apache-2.0 | 2.0 M | from 61.2 GB |
| 159 | Qwen2.5-VL-7B-Instruct-AWQ | 8.3 B | 128 K | ✓ apache-2.0 | 2.0 M | from 9.4 GB |
| 160 | Qwen3-ASR-1.7B | 2.4 B | — | ✓ apache-2.0 | 2.0 M | from 6.0 GB |
| 161 | whisper-tiny | 40 M | — | ✓ apache-2.0 | 2.0 M | from 0.7 GB |
| 162 | Qwen2.5-Coder-7B-Instruct | 7.6 B | 33 K | ✓ apache-2.0 | 2.0 M | from 18.4 GB |
| 163 | wav2vec2-large-xlsr-53-dutch | — | — | ✓ apache-2.0 | 2.0 M | — |
| 164 | SmolLM2-135M-Instruct | 130 M | 8 K | ✓ apache-2.0 | 2.0 M | from 0.8 GB |
| 165 | multi-qa-mpnet-base-dot-v1 | 110 M | 512 | unknown | 1.9 M | from 1.0 GB |
| 166 | stanford-deidentifier-base | — | 512 | ✓ mit | 1.9 M | — |
| 167 | Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive | — | — | ✓ apache-2.0 | 1.9 M | from 13.3 GB |
| 168 | Qwen3-30B-A3B-Instruct-2507 | 30.5 B | 262 K | ✓ apache-2.0 | 1.9 M | from 72.3 GB |
| 169 | Qwen3.6-27B-NVFP4 | 18.2 B | — | ✓ apache-2.0 | 1.9 M | from 27.3 GB |
| 170 | resnet18.a1_in1k | 10 M | — | ✓ apache-2.0 | 1.9 M | from 0.6 GB |
| 171 | parakeet-unified-en-0.6b-gguf | — | — | ✓ cc-by-4.0 | 1.9 M | from 1.0 GB |
| 172 | parakeet-tdt-0.6b-v2 | 620 M | — | ✓ cc-by-4.0 | 1.9 M | from 3.3 GB |
| 173 | parakeet-ctc-1.1b | 1.1 B | — | ✓ cc-by-4.0 | 1.9 M | from 2.0 GB |
| 174 | PowerMoE-3b | 3.4 B | 4 K | ✓ apache-2.0 | 1.8 M | from 15.9 GB |
| 175 | Qwen3.6-35B-A3B-NVFP4 | 24.6 B | — | ✓ apache-2.0 | 1.8 M | from 33.3 GB |
| 176 | SmolLM2-135M | 130 M | 8 K | ✓ apache-2.0 | 1.8 M | from 0.8 GB |
| 177 | gemma-3-4b-it | 4.3 B | — | ⚠ gemma | 1.8 M | from 10.6 GB |
| 178 | bge-base-en-v1.5-course-recommender-v5 | 110 M | 512 | unknown | 1.8 M | from 1.0 GB |
| 179 | ko-sroberta-multitask | 110 M | 512 | unknown | 1.8 M | from 1.0 GB |
| 180 | Meta-Llama-3-8B-Instruct | 8.0 B | — | ⚠ llama3 | 1.8 M | from 19.4 GB |
| 181 | GLM-5.2-NVFP4 | 381.0 B | 1.0 M | ✓ mit | 1.8 M | from 569.0 GB |
| 182 | Qwen2-VL-7B-Instruct-AWQ | 8.3 B | 33 K | ✓ apache-2.0 | 1.8 M | from 9.4 GB |
| 183 | Qwen3-VL-235B-A22B-Instruct | 235.7 B | — | ✓ apache-2.0 | 1.8 M | from 554.3 GB |
| 184 | wav2vec2-indonesian-javanese-sundanese | — | — | ✓ apache-2.0 | 1.8 M | — |
| 185 | bge-large-zh-v1.5 | — | 512 | ✓ mit | 1.8 M | — |
| 186 | Llama-3.2-1B | 1.2 B | — | ⚠ llama3.2 | 1.8 M | from 3.4 GB |
| 187 | Qwen2.5-32B-Instruct | 32.8 B | 33 K | ✓ apache-2.0 | 1.7 M | from 77.5 GB |
| 188 | wav2vec2-base-960h | 90 M | — | ✓ apache-2.0 | 1.7 M | from 0.9 GB |
| 189 | Qwen2.5-Coder-32B-Instruct-AWQ | 32.8 B | 33 K | ✓ apache-2.0 | 1.7 M | from 26.7 GB |
| 190 | stsb-bert-tiny-safetensors | 0 M | 512 | unknown | 1.7 M | from 0.5 GB |
| 191 | gte-large-en-v1.5 | 430 M | 8 K | ✓ apache-2.0 | 1.7 M | from 2.5 GB |
| 192 | diffusiongemma-26B-A4B-it-NVFP4 | 14.4 B | — | ✓ apache-2.0 | 1.7 M | from 23.4 GB |
| 193 | bge-base-zh-v1.5 | — | 512 | ✓ mit | 1.7 M | — |
| 194 | wav2vec2-large-xlsr-53-greek | — | — | ✓ apache-2.0 | 1.7 M | — |
| 195 | distil-large-v3 | 760 M | — | ✓ mit | 1.7 M | from 5.6 GB |
| 196 | DeepSeek-R1-0528-Qwen3-8B | 8.2 B | 131 K | ✓ mit | 1.7 M | from 19.7 GB |
| 197 | tf_efficientnetv2_s.in21k_ft_in1k | 20 M | — | ✓ apache-2.0 | 1.6 M | from 0.6 GB |
| 198 | Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF | — | — | ✓ apache-2.0 | 1.6 M | from 13.8 GB |
| 199 | t5-base | 220 M | — | ✓ apache-2.0 | 1.6 M | from 1.5 GB |
| 200 | SapBERT-from-PubMedBERT-fulltext | 110 M | 512 | ✓ apache-2.0 | 1.6 M | from 1.0 GB |
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.