1,000 models · refreshed nightly
Text generation models
Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.
| # | Model | Params | Context | Commercial use | 30d | Min VRAM |
|---|---|---|---|---|---|---|
| 201 | Laguna-S-2.1-NVFP4 | 117.6 B | 1.0 M | openmdw-1.1 | 436 K | from 127.8 GB |
| 202 | DeepSeek-R1-Distill-Qwen-1.5B | 1.8 B | 131 K | ✓ mit | 431 K | from 4.7 GB |
| 203 | Qwen3-4B-GGUF | — | — | ✓ apache-2.0 | 428 K | from 3.2 GB |
| 204 | Qwen2.5-3B-Instruct-AWQ | 3.4 B | 33 K | other | 427 K | from 4.0 GB |
| 205 | Ornith-1.0-35B-FP8 | 35.1 B | — | ✓ mit | 425 K | from 47.1 GB |
| 206 | maple-preview-GGUF | — | — | ✓ mit | 424 K | from 6.5 GB |
| 207 | Qwen-AgentWorld-35B-A3B-GGUF | — | — | ✓ apache-2.0 | 423 K | from 13.1 GB |
| 208 | MiniCPM5-2B | 2.5 B | 131 K | ✓ apache-2.0 | 421 K | from 6.4 GB |
| 209 | Qwen3-4B-Thinking-2507 | 4.0 B | 262 K | ✓ apache-2.0 | 420 K | from 10.0 GB |
| 210 | Qwen3.6-14B-A3B-FableVibes-GGUF | — | — | ✓ apache-2.0 | 419 K | from 6.4 GB |
| 211 | Mistral-7B-v0.1 | 7.2 B | 33 K | ✓ apache-2.0 | 415 K | from 17.5 GB |
| 212 | saiga_llama3_8b | 8.0 B | 8 K | other | 414 K | from 19.4 GB |
| 213 | GPT-OSS-20B-NPU2 | 20.0 B | 131 K | ✓ apache-2.0 | 411 K | from 47.5 GB |
| 214 | Ornith-1.5-35B-A3B | 36.0 B | — | ✓ mit | 410 K | from 85.0 GB |
| 215 | NVIDIA-Nemotron-3-Nano-30B-A3B-FP8 | 31.6 B | 262 K | other | 409 K | from 41.2 GB |
| 216 | Phi-3-mini-4k-instruct | 3.8 B | 4 K | ✓ mit | 407 K | from 9.5 GB |
| 217 | Qwen3-8B-GGUF | — | — | ✓ apache-2.0 | 404 K | from 6.0 GB |
| 218 | gemma-2-9b-it-AWQ-INT4 | 9.2 B | 8 K | ⚠ gemma | 402 K | from 8.7 GB |
| 219 | Qwen2.5-VL-7B-Instruct-NVFP4 | 5.0 B | 128 K | other | 399 K | from 9.2 GB |
| 220 | EXAONE-3.5-7.8B-Instruct | 7.8 B | 33 K | other | 396 K | from 36.1 GB |
| 221 | gemma-3-270m | 270 M | — | ⚠ gemma | 395 K | from 1.1 GB |
| 222 | granite-4.1-3b | 3.4 B | 131 K | ✓ apache-2.0 | 393 K | from 8.5 GB |
| 223 | Ternary-Bonsai-8B-gguf | — | — | ✓ apache-2.0 | 392 K | from 2.9 GB |
| 224 | Qwen3.8-27B-DFlash2 | 1.9 B | 262 K | ✓ apache-2.0 | 391 K | from 5.0 GB |
| 225 | Kwaipilot_KAT-Coder-V2.5-Dev-GGUF | — | — | ✓ apache-2.0 | 390 K | from 11.3 GB |
| 226 | Ornith-1.0-397B | 396.8 B | — | ✓ mit | 387 K | from 933.0 GB |
| 227 | Phi-4-mini-instruct | 3.8 B | 131 K | ✓ mit | 384 K | from 9.5 GB |
| 228 | TwIL-LM3 | 3.1 B | 66 K | other | 384 K | from 3.1 GB |
| 229 | NVIDIA-Nemotron-Nano-9B-v2 | 8.9 B | 131 K | other | 383 K | from 21.4 GB |
| 230 | Ornith-1.5-35B-A3B-MLX | 34.7 B | — | unknown | 383 K | from 82.0 GB |
| 231 | Ornith-1.5-397B-FP8 | 403.4 B | — | ✓ mit | 382 K | from 521.2 GB |
| 232 | falcon-7b | 7.2 B | — | ✓ apache-2.0 | 379 K | from 17.5 GB |
| 233 | pythia-14m | 10 M | 2 K | ✓ apache-2.0 | 378 K | from 0.5 GB |
| 234 | Hermes-3-Llama-3.1-8B | 8.0 B | 131 K | ⚠ llama3 | 377 K | from 19.4 GB |
| 235 | mistral-7b-v0.3-bnb-4bit | 7.5 B | 33 K | ✓ apache-2.0 | 377 K | from 6.2 GB |
| 236 | Qwen3-235B-A22B-Instruct-2507-FP8 | 235.1 B | 262 K | ✓ apache-2.0 | 376 K | from 295.8 GB |
| 237 | Ornith-1.5-9B-OBLITERATED | 9.7 B | — | ✓ mit | 371 K | from 6.3 GB |
| 238 | Qwen2-7B-Instruct | 7.6 B | 33 K | ✓ apache-2.0 | 369 K | from 18.4 GB |
| 239 | Llama-3.2-3B | 3.2 B | — | ⚠ llama3.2 | 368 K | from 8.0 GB |
| 240 | Parable-Granite-4.1-3B-Claude-Fable-5-GGUF | — | — | ✓ apache-2.0 | 366 K | from 2.8 GB |
| 241 | Qwen3.8-27B-DFlash2 | 1.9 B | 262 K | ✓ apache-2.0 | 364 K | from 5.0 GB |
| 242 | Parable-Granite-4.1-8B-Claude-Fable-5-GGUF | — | — | ✓ apache-2.0 | 362 K | from 6.1 GB |
| 243 | Qwen2.5-Coder-1.5B | 1.5 B | 33 K | ✓ apache-2.0 | 361 K | from 4.1 GB |
| 244 | MiniCPM5-1B-Claude-Opus-Fable5-V2-Thinking-GGUF | — | — | ✓ apache-2.0 | 360 K | from 1.8 GB |
| 245 | Qwen3.6-35B-A3B-NVFP4-MTP-GGUF | — | — | unknown | 360 K | from 23.0 GB |
| 246 | POCKET-26B-GGUF | — | — | ✓ apache-2.0 | 360 K | from 12.7 GB |
| 247 | NVIDIA-Nemotron-3.5-Lightning-30B-A3B-FP8 | 31.6 B | 262 K | unknown | 357 K | from 40.8 GB |
| 248 | Qwen2.5-72B-Instruct | 72.7 B | 33 K | other | 356 K | from 171.4 GB |
| 249 | vlt5-base-keywords | 280 M | — | ✓ cc-by-4.0 | 355 K | from 1.8 GB |
| 250 | Ornith-1.0-397B-FP8 | 396.8 B | — | ✓ mit | 354 K | from 505.7 GB |
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.