1,060 models · refreshed nightly
All models
Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.
| # | Model | Params | Context | Commercial use | 30d | Min VRAM |
|---|---|---|---|---|---|---|
| 801 | Meta-Llama-3.1-8B-Instruct | 8.0 B | 131 K | ⚠ llama3.1 | 243 K | from 19.4 GB |
| 802 | gliner_multi-v2.1 | 290 M | — | ✓ apache-2.0 | 242 K | from 1.8 GB |
| 803 | privacy-filter | 1.4 B | 131 K | ✓ apache-2.0 | 242 K | from 6.9 GB |
| 804 | h2ovl-mississippi-800m | 830 M | — | ✓ apache-2.0 | 242 K | from 2.4 GB |
| 805 | SmolLM2-1.7B | 1.7 B | 8 K | ✓ apache-2.0 | 241 K | from 4.5 GB |
| 806 | convnext_femto.d1_in1k | 10 M | — | ✓ apache-2.0 | 241 K | from 0.5 GB |
| 807 | Qwen3-8B-FP8 | 8.2 B | 41 K | ✓ apache-2.0 | 240 K | from 12.1 GB |
| 808 | pythia-410m | 510 M | 2 K | ✓ apache-2.0 | 239 K | from 1.6 GB |
| 809 | Vikhr-Nemo-12B-Instruct-R-21-09-24 | 12.3 B | 1.0 M | ✓ apache-2.0 | 239 K | from 30.8 GB |
| 810 | segformer_b2_clothes | 30 M | — | other | 238 K | from 0.6 GB |
| 811 | gemma-3-270m | 270 M | — | ⚠ gemma | 238 K | from 1.1 GB |
| 812 | NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4-DSpark | 760 M | 1.0 M | other | 237 K | from 2.1 GB |
| 813 | deberta-xlarge-mnli | — | 512 | ✓ mit | 237 K | — |
| 814 | span-marker-bert-base-uncased-acronyms | 110 M | — | ✓ apache-2.0 | 237 K | from 1.0 GB |
| 815 | Qwen3.6-35B-A3B-Claude-4.6-Opus-Reasoning-Distilled-GGUF | — | — | ✓ apache-2.0 | 236 K | from 23.8 GB |
| 816 | Qwen-Image-2512-GGUF | — | — | ✓ apache-2.0 | 236 K | from 8.6 GB |
| 817 | Mellum2-12B-A2.5B-Instruct-GGUF-Q8_0 | — | — | ✓ apache-2.0 | 236 K | from 14.7 GB |
| 818 | roberta_toxicity_classifier | — | 512 | ⚠ openrail++ | 235 K | — |
| 819 | Laguna-XS-2.1-GGUF | — | — | openmdw-1.1 | 234 K | from 22.8 GB |
| 820 | detr-doc-table-detection | 40 M | 1 K | ✓ apache-2.0 | 234 K | from 0.7 GB |
| 821 | Qwen3-0.6B-GGUF | 600 M | 41 K | ✓ apache-2.0 | 233 K | from 1.3 GB |
| 822 | ner-english | — | — | unknown | 233 K | — |
| 823 | fullstop-punctuation-multilang-large | 560 M | 512 | ✓ mit | 232 K | from 3.0 GB |
| 824 | diffusiongemma-26B-A4B-it-NVFP4 | 14.4 B | — | ✓ apache-2.0 | 232 K | from 23.4 GB |
| 825 | Qwen2.5-14B-Instruct-GPTQ-Int4 | 14.8 B | 33 K | ✓ apache-2.0 | 232 K | from 13.7 GB |
| 826 | Mistral-Small-24B-Instruct-2501-AWQ | 23.6 B | 33 K | ✓ apache-2.0 | 232 K | from 19.7 GB |
| 827 | biomedical-ner-all | 66 M | 512 | ✓ apache-2.0 | 231 K | from 0.8 GB |
| 828 | llama-7b | 6.7 B | 2 K | other | 231 K | from 16.3 GB |
| 829 | gliner2.5-multi-v1 | 287 M | — | ✓ apache-2.0 | 231 K | from 1.8 GB |
| 830 | h2ovl-mississippi-2b | 2.2 B | — | ✓ apache-2.0 | 230 K | from 5.6 GB |
| 831 | Qwen2.5-1.5B-Instruct-GGUF | — | — | ✓ apache-2.0 | 230 K | from 1.3 GB |
| 832 | pegasus-xsum | — | 512 | unknown | 230 K | — |
| 833 | Agents-A1-4B-Q8_0-GGUF | — | — | ✓ apache-2.0 | 229 K | from 1.2 GB |
| 834 | twitter-xlm-roberta-base-sentiment-multilingual | — | 512 | unknown | 229 K | — |
| 835 | Ornith-1.0-397B | 396.8 B | — | ✓ mit | 228 K | from 933.0 GB |
| 836 | OpenMed-NER-ChemicalDetect-ModernMed-149M | 150 M | 8 K | ✓ apache-2.0 | 225 K | from 0.9 GB |
| 837 | TinyStories-1M | 1.0 M | 2 K | unknown | 225 K | from 0.5 GB |
| 838 | Meta-Llama-3.1-8B-Instruct-AWQ-INT4 | 8.0 B | 131 K | ⚠ llama3.1 | 224 K | from 8.0 GB |
| 839 | distilbert-imdb | — | 512 | ✓ apache-2.0 | 224 K | — |
| 840 | BiRefNet_lite | 44 M | — | ✓ mit | 223 K | from 0.7 GB |
| 841 | Yi-Coder-9B-Chat-GGUF | — | — | unknown | 223 K | from 2.7 GB |
| 842 | Kimi-K2-Instruct | 1,026.5 B | 131 K | other | 222 K | from 1,286.6 GB |
| 843 | MiniMax-M2.5-NVFP4 | 116.4 B | 197 K | other | 222 K | from 171.8 GB |
| 844 | Ornith-1.5-9B-MLX-6bit | 9.0 B | — | unknown | 221 K | from 9.8 GB |
| 845 | hyenadna-medium-450k-seqlen-hf | 28 M | — | ✓ bsd-3-clause | 221 K | from 0.6 GB |
| 846 | GLM-4.7-Flash-AWQ-4bit | 32.1 B | 203 K | ✓ mit | 221 K | from 27.5 GB |
| 847 | LLaDA2.0-mini | 16.3 B | 33 K | ✓ apache-2.0 | 220 K | from 38.7 GB |
| 848 | Olmo-3-7B-Instruct-SFT | 7.3 B | 66 K | ✓ apache-2.0 | 220 K | from 17.7 GB |
| 849 | opus-mt-es-en | — | 512 | ✓ apache-2.0 | 220 K | — |
| 850 | automotive | 34.7 B | — | ✓ apache-2.0 | 219 K | from 29.0 GB |
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.