1,000 models · refreshed nightly
All models
Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.
| # | Model | Params | Context | Commercial use | 30d | Min VRAM |
|---|---|---|---|---|---|---|
| 601 | Ternary-Bonsai-8B-gguf | — | — | ✓ apache-2.0 | 392 K | from 2.9 GB |
| 602 | Qwen3.8-27B-DFlash2 | 1.9 B | 262 K | ✓ apache-2.0 | 391 K | from 5.0 GB |
| 603 | Kwaipilot_KAT-Coder-V2.5-Dev-GGUF | — | — | ✓ apache-2.0 | 390 K | from 11.3 GB |
| 604 | Ornith-1.0-397B | 396.8 B | — | ✓ mit | 387 K | from 933.0 GB |
| 605 | Phi-4-mini-instruct | 3.8 B | 131 K | ✓ mit | 384 K | from 9.5 GB |
| 606 | TwIL-LM3 | 3.1 B | 66 K | other | 384 K | from 3.1 GB |
| 607 | NVIDIA-Nemotron-Nano-9B-v2 | 8.9 B | 131 K | other | 383 K | from 21.4 GB |
| 608 | Ornith-1.5-35B-A3B-MLX | 34.7 B | — | unknown | 383 K | from 82.0 GB |
| 609 | Ornith-1.5-397B-FP8 | 403.4 B | — | ✓ mit | 382 K | from 521.2 GB |
| 610 | rtdetr_r50vd_coco_o365 | 40 M | — | ✓ apache-2.0 | 380 K | from 0.7 GB |
| 611 | falcon-7b | 7.2 B | — | ✓ apache-2.0 | 379 K | from 17.5 GB |
| 612 | pythia-14m | 10 M | 2 K | ✓ apache-2.0 | 378 K | from 0.5 GB |
| 613 | Hermes-3-Llama-3.1-8B | 8.0 B | 131 K | ⚠ llama3 | 377 K | from 19.4 GB |
| 614 | mistral-7b-v0.3-bnb-4bit | 7.5 B | 33 K | ✓ apache-2.0 | 377 K | from 6.2 GB |
| 615 | Qwen3-235B-A22B-Instruct-2507-FP8 | 235.1 B | 262 K | ✓ apache-2.0 | 376 K | from 295.8 GB |
| 616 | opus-mt-ko-en | — | 512 | ✓ apache-2.0 | 372 K | — |
| 617 | Ornith-1.5-9B-OBLITERATED | 9.7 B | — | ✓ mit | 371 K | from 6.3 GB |
| 618 | Qwen2-7B-Instruct | 7.6 B | 33 K | ✓ apache-2.0 | 369 K | from 18.4 GB |
| 619 | segformer-b0-finetuned-ade-512-512 | <0.1 M | — | other | 368 K | from 0.5 GB |
| 620 | Llama-3.2-3B | 3.2 B | — | ⚠ llama3.2 | 368 K | from 8.0 GB |
| 621 | sanskrit-en-custom-transformer | — | — | ✓ mit | 367 K | — |
| 622 | Parable-Granite-4.1-3B-Claude-Fable-5-GGUF | — | — | ✓ apache-2.0 | 366 K | from 2.8 GB |
| 623 | privacy-filter-multilingual-GGUF | — | — | ✓ apache-2.0 | 366 K | from 2.3 GB |
| 624 | Qwen3.8-27B-DFlash2 | 1.9 B | 262 K | ✓ apache-2.0 | 364 K | from 5.0 GB |
| 625 | Parable-Granite-4.1-8B-Claude-Fable-5-GGUF | — | — | ✓ apache-2.0 | 362 K | from 6.1 GB |
| 626 | Qwen2.5-Coder-1.5B | 1.5 B | 33 K | ✓ apache-2.0 | 361 K | from 4.1 GB |
| 627 | MiniCPM5-1B-Claude-Opus-Fable5-V2-Thinking-GGUF | — | — | ✓ apache-2.0 | 360 K | from 1.8 GB |
| 628 | Qwen3.6-35B-A3B-NVFP4-MTP-GGUF | — | — | unknown | 360 K | from 23.0 GB |
| 629 | POCKET-26B-GGUF | — | — | ✓ apache-2.0 | 360 K | from 12.7 GB |
| 630 | NVIDIA-Nemotron-3.5-Lightning-30B-A3B-FP8 | 31.6 B | 262 K | unknown | 357 K | from 40.8 GB |
| 631 | Qwen2.5-72B-Instruct | 72.7 B | 33 K | other | 356 K | from 171.4 GB |
| 632 | Z-Image-Turbo-GGUF | — | — | ✓ apache-2.0 | 356 K | from 4.5 GB |
| 633 | vlt5-base-keywords | 280 M | — | ✓ cc-by-4.0 | 355 K | from 1.8 GB |
| 634 | Ornith-1.0-397B-FP8 | 396.8 B | — | ✓ mit | 354 K | from 505.7 GB |
| 635 | AI-image-detector | — | — | ✓ cc-by-4.0 | 353 K | — |
| 636 | GLM-5.2-GGUF | — | — | ✓ mit | 352 K | from 52.0 GB |
| 637 | t5-large | 740 M | — | ✓ apache-2.0 | 352 K | from 3.9 GB |
| 638 | Z-Image-Lora | — | — | ✓ apache-2.0 | 351 K | from 109.6 GB |
| 639 | amd.Instella-MoE-16B-A3B-Think-GGUF | — | — | unknown | 351 K | from 7.7 GB |
| 640 | deberta-xlarge-mnli | — | 512 | ✓ mit | 349 K | — |
| 641 | RMBG-1.4 | 40 M | — | other | 349 K | from 0.7 GB |
| 642 | Qwen3-235B-A22B | 235.1 B | 41 K | ✓ apache-2.0 | 348 K | from 553.0 GB |
| 643 | wide_resnet50_2.racm_in1k | 70 M | — | ✓ apache-2.0 | 348 K | from 0.8 GB |
| 644 | animagine-xl-4.0 | 2.6 B | — | ⚠ openrail++ | 344 K | from 23.8 GB |
| 645 | Agents-A1-4B | 4.5 B | — | ✓ apache-2.0 | 344 K | from 11.2 GB |
| 646 | NVIDIA-Nemotron-3-Ultra-550B-A55B-NVFP4 | 302.8 B | 262 K | other | 343 K | from 433.5 GB |
| 647 | Ling-3.0-flash-GGUF | — | — | ✓ mit | 341 K | from 36.1 GB |
| 648 | Ornith-1.5-9B-MLX-8bit | 9.0 B | — | unknown | 341 K | from 12.3 GB |
| 649 | Ornith-1.5-9B-MLX | 9.0 B | — | unknown | 341 K | from 21.5 GB |
| 650 | RealVisXL_V5.0 | 2.6 B | — | ⚠ openrail++ | 341 K | from 46.7 GB |
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.