1,011 models · refreshed nightly

All models

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dMin VRAM
51 speaker-diarization-community-1 pyannote · ASR · gated ✓ cc-by-4.0 5.3 M
52 nomic-embed-text-v1 nomic-ai · sentence-similarity 140 M 8 K ✓ apache-2.0 5.1 M from 1.1 GB
53 Ornith-1.0-9B-GGUF deepreinforce-ai · text-generation ✓ mit 4.9 M from 6.7 GB
54 Ornith-1.0-9B-GGUF ornith-ai · text-generation ✓ mit 4.9 M from 6.7 GB
55 mxbai-embed-large-v1 mixedbread-ai · feature-extraction 340 M 512 ✓ apache-2.0 4.9 M from 1.3 GB
56 Qwen3-VL-8B-Instruct Qwen · image-text-to-text 8.8 B ✓ apache-2.0 4.9 M from 21.1 GB
57 whisper-base openai · ASR 70 M ✓ apache-2.0 4.8 M from 0.8 GB
58 gemma-4-31B-it-FP8-block RedHatAI · image-text-to-text 31.3 B ✓ apache-2.0 4.8 M from 41.8 GB
59 bge-small-zh-v1.5 BAAI · feature-extraction 20 M 512 ✓ mit 4.8 M from 0.6 GB
60 gemma-3-1b-it google · text-generation · gated 1.0 B ⚠ gemma 4.7 M from 2.8 GB
61 Qwen3-Coder-30B-A3B-Instruct-GGUF unsloth · text-generation ✓ apache-2.0 4.7 M from 12.9 GB
62 vit-base-patch16-224 google · image-classification 90 M ✓ apache-2.0 4.7 M from 0.9 GB
63 OTel-LLM-E4B-IT farbodtavakkoli · text-generation ✓ apache-2.0 4.7 M
64 wav2vec2-large-xlsr-53-portuguese jonatasgrosman · ASR ✓ apache-2.0 4.6 M
65 Qwen2.5-7B-Instruct-AWQ Qwen · text-generation 7.6 B 33 K ✓ apache-2.0 4.6 M from 7.8 GB
66 dolphin-2.9.1-yi-1.5-34b dphn · text-generation 34.4 B 8 K ✓ apache-2.0 4.6 M from 81.3 GB
67 Qwen3-4B Qwen · text-generation 4.0 B 41 K ✓ apache-2.0 4.6 M from 10.0 GB
68 bge-reranker-base BAAI · text-classification 280 M 512 ✓ mit 4.6 M from 1.8 GB
69 Prompt-Guard-86M meta-llama · text-classification · gated 280 M ⚠ llama3.1 4.5 M from 1.8 GB
70 Qwen3-ASR-0.6B Qwen · ASR 940 M ✓ apache-2.0 4.3 M from 2.7 GB
71 gpt-oss-120b openai · text-generation 120.4 B 131 K ✓ apache-2.0 4.2 M from 162.1 GB
72 granite-4.1-8b ibm-granite · text-generation 8.8 B 131 K ✓ apache-2.0 4.2 M from 21.2 GB
73 wav2vec2-large-xlsr-53-russian jonatasgrosman · ASR ✓ apache-2.0 4.0 M
74 Qwen3-VL-4B-Instruct Qwen · image-text-to-text 4.4 B ✓ apache-2.0 4.0 M from 10.9 GB
75 granite-embedding-small-english-r2 ibm-granite · feature-extraction 50 M 8 K ✓ apache-2.0 3.8 M from 0.6 GB
76 GLM-OCR zai-org · image-text-to-text 1.3 B ✓ mit 3.8 M from 3.6 GB
77 Ornith-1.0-35B-GGUF deepreinforce-ai · text-generation ✓ mit 3.8 M from 23.8 GB
78 Ornith-1.0-35B-GGUF ornith-ai · text-generation ✓ mit 3.8 M from 23.8 GB
79 gemma-4-26B-A4B-it-AWQ-4bit cyankiwi · image-text-to-text 26.6 B ✓ apache-2.0 3.7 M from 23.4 GB
80 distilbert-base-uncased-finetuned-sst-2-english distilbert · text-classification 70 M 512 ✓ apache-2.0 3.7 M from 0.8 GB
81 all-MiniLM-L12-v2 sentence-transformers · sentence-similarity 30 M 512 ✓ apache-2.0 3.5 M from 0.7 GB
82 Qwen3.6-27B-NVFP4 unsloth · image-text-to-text 21.2 B ✓ apache-2.0 3.4 M from 29.4 GB
83 Qwen3-Embedding-4B Qwen · feature-extraction 4.0 B 41 K ✓ apache-2.0 3.4 M from 10.0 GB
84 Qwen2.5-14B-Instruct Qwen · text-generation 14.8 B 33 K ✓ apache-2.0 3.3 M from 35.2 GB
85 Qwen3-Embedding-8B Qwen · feature-extraction 7.6 B 41 K ✓ apache-2.0 3.3 M from 18.3 GB
86 OTel-2.0-LLM-31B-IT farbodtavakkoli · text-generation 32.1 B 262 K ✓ apache-2.0 3.3 M from 76.0 GB
87 jina-embeddings-v3 jinaai · feature-extraction 570 M 8 K ✗ cc-by-nc-4.0 3.2 M from 1.8 GB
88 Qwen3-4B-Instruct-2507 Qwen · text-generation 4.0 B 262 K ✓ apache-2.0 3.2 M from 10.0 GB
89 Qwen3-VL-8B-Instruct-FP8 Qwen · image-text-to-text 8.8 B ✓ apache-2.0 3.2 M from 13.5 GB
90 fairface_age_image_detection dima806 · image-classification 90 M ✓ apache-2.0 3.2 M from 2.4 GB
91 Qwen2-1.5B-Instruct Qwen · text-generation 1.5 B 33 K ✓ apache-2.0 3.2 M from 4.1 GB
92 chandra-ocr-2 datalab-to · image-text-to-text 5.3 B ⚠ openrail 3.2 M from 12.9 GB
93 llava-1.5-7b-hf llava-hf · image-text-to-text 7.1 B ⚠ llama2 3.1 M from 17.1 GB
94 indonesian-roberta-base-posp-tagger w11wo · token-classification 120 M 512 ✓ mit 3.1 M from 1.1 GB
95 Qwen3.5-0.8B Qwen · image-text-to-text 870 M ✓ apache-2.0 3.1 M from 2.6 GB
96 Qwen3-30B-A3B Qwen · text-generation 30.5 B 41 K ✓ apache-2.0 3.1 M from 72.3 GB
97 Qwen2-VL-2B-Instruct Qwen · image-text-to-text 2.2 B 33 K ✓ apache-2.0 3.0 M from 5.7 GB
98 Qwen2.5-Coder-14B-Instruct Qwen · text-generation 14.8 B 33 K ✓ apache-2.0 3.0 M from 35.2 GB
99 all-MiniLM-L6-v2 Xenova · feature-extraction 512 ✓ apache-2.0 2.9 M
100 koelectra-small-v3-nsmc daekeun-ml · text-classification 10 M 512 ✓ mit 2.9 M from 0.6 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.