1,000 models · refreshed nightly
All models
Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.
| # | Model | Params | Context | Commercial use | 30d | Min VRAM |
|---|---|---|---|---|---|---|
| 451 | DeepSeek-R1-Distill-Qwen-1.5B | 1.8 B | 131 K | ✓ mit | 603 K | from 4.7 GB |
| 452 | bloom-560m | 560 M | — | ⚠ bigscience-bloom-rail-1.0 | 599 K | from 1.8 GB |
| 453 | mask2former-swin-large-ade-semantic | 220 M | — | other | 598 K | from 1.5 GB |
| 454 | Phi-3-mini-4k-instruct | 3.8 B | 4 K | ✓ mit | 597 K | from 9.5 GB |
| 455 | deepseek-coder-7b-instruct-v1.5 | 6.9 B | 4 K | other | 594 K | from 16.7 GB |
| 456 | Bonsai-27B-mlx-1bit | 1.7 B | — | ✓ apache-2.0 | 594 K | from 6.4 GB |
| 457 | Ternary-Bonsai-27B-mlx-2bit | 2.6 B | — | ✓ apache-2.0 | 592 K | from 10.2 GB |
| 458 | cryptobert | 120 M | 512 | ✓ mit | 589 K | from 1.1 GB |
| 459 | conv-bert-base | — | 512 | unknown | 589 K | — |
| 460 | NVIDIA-Nemotron-3-Nano-30B-A3B-FP8 | 31.6 B | 262 K | other | 587 K | from 41.2 GB |
| 461 | Qwen2.5-Coder-7B-Instruct-AWQ | 7.6 B | 33 K | ✓ apache-2.0 | 586 K | from 7.9 GB |
| 462 | Qwen3-30B-A3B-Instruct-2507-FP8 | 30.5 B | 262 K | ✓ apache-2.0 | 582 K | from 39.4 GB |
| 463 | snowflake-arctic-embed-s | 30 M | 512 | ✓ apache-2.0 | 581 K | from 0.7 GB |
| 464 | nb-wav2vec2-1b-bokmaal-v2 | 960 M | — | ✓ apache-2.0 | 580 K | from 4.9 GB |
| 465 | encodec_24khz | 20 M | — | unknown | 577 K | from 0.6 GB |
| 466 | LFM2.5-1.2B-Instruct | 1.2 B | 128 K | other | 575 K | from 3.3 GB |
| 467 | nsfw_image_detector | 90 M | — | ✓ mit | 575 K | from 0.7 GB |
| 468 | sentence-bert-base-ja-mean-tokens-v2 | 110 M | 512 | cc-by-sa-4.0 | 573 K | from 1.0 GB |
| 469 | Phi-4-multimodal-instruct | 5.6 B | 131 K | ✓ mit | 573 K | from 15.4 GB |
| 470 | Qwen2-7B-Instruct | 7.6 B | 33 K | ✓ apache-2.0 | 569 K | from 18.4 GB |
| 471 | gpt2-medium | 380 M | — | ✓ mit | 565 K | from 2.2 GB |
| 472 | VLM2Vec-Full | 4.2 B | 131 K | ✓ apache-2.0 | 563 K | from 10.2 GB |
| 473 | wav2vec2-xls-r-parlaspeech-hr | 320 M | — | unknown | 562 K | from 1.9 GB |
| 474 | DeepSeek-V4-Flash-DSpark | 165.3 B | 1.0 M | ✓ mit | 561 K | from 208.9 GB |
| 475 | Qwen3-Reranker-4B-W4A16-G128 | 4.1 B | 41 K | ✓ apache-2.0 | 561 K | from 4.0 GB |
| 476 | MERT-v1-330M | — | — | ✗ cc-by-nc-4.0 | 559 K | — |
| 477 | MiniMax-M3-NVFP4 | 246.6 B | — | other | 551 K | from 312.6 GB |
| 478 | DeepSeek-Coder-V2-Lite-Instruct | 15.7 B | 164 K | other | 546 K | from 37.4 GB |
| 479 | msmarco-bert-base-dot-v5 | 110 M | 512 | unknown | 544 K | from 1.0 GB |
| 480 | deepseek-coder-6.7b-instruct | 6.7 B | 16 K | other | 542 K | from 16.3 GB |
| 481 | Hermes-3-Llama-3.1-8B | 8.0 B | 131 K | ⚠ llama3 | 542 K | from 19.4 GB |
| 482 | gpt-oss-20b-GGUF | — | 131 K | ✓ apache-2.0 | 542 K | from 13.1 GB |
| 483 | wav2vec2-xlsr-nepali | — | — | ✓ apache-2.0 | 542 K | — |
| 484 | sat-3l-sm | 210 M | 514 | ✓ mit | 541 K | from 1.5 GB |
| 485 | ner-english-fast | — | — | unknown | 539 K | — |
| 486 | detr-resnet-50 | 40 M | 1 K | ✓ apache-2.0 | 538 K | from 0.7 GB |
| 487 | Qwen3-Coder-Next | 79.7 B | 262 K | ✓ apache-2.0 | 537 K | from 187.7 GB |
| 488 | gemma-2-9b-it-AWQ-INT4 | 9.2 B | 8 K | ⚠ gemma | 535 K | from 8.7 GB |
| 489 | Qwen3-TTS-12Hz-1.7B-VoiceDesign | 1.9 B | — | ✓ apache-2.0 | 533 K | from 5.8 GB |
| 490 | Llama-3.2-3B | 3.2 B | — | ⚠ llama3.2 | 533 K | from 8.0 GB |
| 491 | macbert4csc-base-chinese | 100 M | 512 | ✓ apache-2.0 | 531 K | from 1.0 GB |
| 492 | whisper-bemba-stt | 240 M | — | unknown | 530 K | from 1.6 GB |
| 493 | gte-base | 110 M | 512 | ✓ mit | 527 K | from 0.8 GB |
| 494 | twitter-roberta-base-sentiment | — | 512 | unknown | 527 K | — |
| 495 | Qwen2.5-Coder-7B-Instruct-GPTQ-Int4 | 7.6 B | 33 K | ✓ apache-2.0 | 527 K | from 7.8 GB |
| 496 | privacy-filter | 1.4 B | 131 K | ✓ apache-2.0 | 525 K | from 6.9 GB |
| 497 | FLUX.1-dev | 11.9 B | — | other | 523 K | from 66.0 GB |
| 498 | EXAONE-3.5-7.8B-Instruct | 7.8 B | 33 K | other | 516 K | from 36.1 GB |
| 499 | jina-embeddings-v5-text-nano | 210 M | 8 K | ✗ cc-by-nc-4.0 | 515 K | from 1.1 GB |
| 500 | Qwen3-Embedding-4B-W4A16-G128 | 4.1 B | 41 K | ✓ apache-2.0 | 514 K | from 4.0 GB |
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.