1,054 models · refreshed nightly

All models

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dMin VRAM
51 Ornith-1.5-9B-GGUF ornith-ai · text-generation ✓ mit 5.8 M from 6.9 GB
52 Qwen3.6-27B-FP8 Qwen · image-text-to-text 27.8 B ✓ apache-2.0 5.8 M from 38.6 GB
53 speaker-diarization-community-1 pyannote · ASR · gated ✓ cc-by-4.0 5.3 M
54 finbert ProsusAI · text-classification 512 unknown 5.3 M
55 gpt-oss-120b openai · text-generation 120.4 B 131 K ✓ apache-2.0 5.1 M from 162.1 GB
56 bge-small-zh-v1.5 BAAI · feature-extraction 20 M 512 ✓ mit 5.1 M from 0.6 GB
57 Qwen2.5-3B-Instruct Qwen · text-generation 3.1 B 33 K other 5.1 M from 7.8 GB
58 Qwen3-32B Qwen · text-generation 32.8 B 41 K ✓ apache-2.0 5.0 M from 77.5 GB
59 Ornith-1.0-9B-GGUF deepreinforce-ai · text-generation ✓ mit 4.9 M from 6.7 GB
60 dolphin-2.9.1-yi-1.5-34b dphn · text-generation 34.4 B 8 K ✓ apache-2.0 4.8 M from 81.3 GB
61 Qwen3.5-2B Qwen · image-text-to-text 2.3 B ✓ apache-2.0 4.8 M from 5.8 GB
62 Qwen3.8-27B-MLX-4bit lmstudio-community · image-text-to-text 4.7 B ✓ apache-2.0 4.7 M from 18.9 GB
63 whisper-large-v3 openai · ASR 1.5 B ✓ apache-2.0 4.7 M from 10.9 GB
64 Ornith-1.5-35B-A3B-GGUF ornith-ai · text-generation ✓ mit 4.6 M from 24.4 GB
65 Qwen3.8-27B-MLX-8bit lmstudio-community · image-text-to-text 8.0 B ✓ apache-2.0 4.5 M from 34.2 GB
66 Qwen3.8-27B-MLX-6bit lmstudio-community · image-text-to-text 6.4 B ✓ apache-2.0 4.5 M from 26.5 GB
67 Qwen3.8-27B-MLX-5bit lmstudio-community · image-text-to-text 5.5 B ✓ apache-2.0 4.5 M from 22.7 GB
68 Prompt-Guard-86M meta-llama · text-classification · gated 280 M ⚠ llama3.1 4.4 M from 1.8 GB
69 audio.cpp-gguf audio-cpp · text-to-speech other 4.3 M from 1.4 GB
70 bge-base-en-v1.5-course-recommender-v5 datasocietyco · sentence-similarity 110 M 512 unknown 4.2 M from 1.0 GB
71 all-MiniLM-L12-v2 sentence-transformers · sentence-similarity 30 M 512 ✓ apache-2.0 4.2 M from 0.7 GB
72 DeepSeek-V4-Flash-0731 deepseek-ai · text-generation 304.2 B 1.0 M ✓ mit 4.2 M from 229.7 GB
73 wav2vec2-large-xlsr-53-russian jonatasgrosman · ASR ✓ apache-2.0 4.2 M
74 Qwen-72B Qwen · text-generation 72.3 B 33 K other 4.2 M from 170.4 GB
75 bge-reranker-base BAAI · text-classification 280 M 512 ✓ mit 4.0 M from 1.8 GB
76 Qwen3-4B-Instruct-2507 Qwen · text-generation 4.0 B 262 K ✓ apache-2.0 4.0 M from 10.0 GB
77 Qwen3.8-27B-iMatrix-NVFP4-MTP-GGUF cdiamond · image-text-to-text ✓ apache-2.0 4.0 M from 19.3 GB
78 Qwen3-1.7B Qwen · text-generation 2.0 B 41 K ✓ apache-2.0 3.8 M from 5.3 GB
79 Ornith-1.0-35B-GGUF deepreinforce-ai · text-generation ✓ mit 3.8 M from 23.8 GB
80 wav2vec2-large-xlsr-53-polish jonatasgrosman · ASR ✓ apache-2.0 3.7 M
81 Qwen3-VL-4B-Instruct Qwen · image-text-to-text 4.4 B ✓ apache-2.0 3.7 M from 10.9 GB
82 distilbert-base-uncased-finetuned-sst-2-english distilbert · text-classification 70 M 512 ✓ apache-2.0 3.7 M from 0.8 GB
83 Qwen3.6-27B Qwen · image-text-to-text 27.8 B ✓ apache-2.0 3.6 M from 65.8 GB
84 Kimi-K3-DSpark RadixArk · text-generation 2.3 B 1.0 M unknown 3.6 M from 5.8 GB
85 Qwen2.5-7B-Instruct-AWQ Qwen · text-generation 7.6 B 33 K ✓ apache-2.0 3.5 M from 7.8 GB
86 NVIDIA-Nemotron-3-Nano-4B-BF16 nvidia · text-generation 4.0 B 262 K other 3.5 M from 9.8 GB
87 pythia-160m EleutherAI · text-generation 210 M 2 K ✓ apache-2.0 3.5 M from 0.9 GB
88 bert-large-cased-finetuned-conll03-english dbmdz · token-classification 330 M 512 unknown 3.4 M from 2.0 GB
89 nsfw_image_detection Falconsai · image-classification 90 M ✓ apache-2.0 3.4 M from 0.9 GB
90 Qwen3.6-35B-A3B Qwen · image-text-to-text 36.0 B ✓ apache-2.0 3.3 M from 85.0 GB
91 Ornith-1.0-9B-GGUF ornith-ai · text-generation ✓ mit 3.3 M from 6.7 GB
92 nomic-embed-text-v1 nomic-ai · sentence-similarity 140 M 8 K ✓ apache-2.0 3.2 M from 1.1 GB
93 twitter-roberta-base-sentiment-latest cardiffnlp · text-classification 512 ✓ cc-by-4.0 3.2 M
94 wav2vec2-large-xlsr-53-dutch jonatasgrosman · ASR ✓ apache-2.0 3.1 M
95 wav2vec2-indonesian-javanese-sundanese indonesian-nlp · ASR ✓ apache-2.0 3.1 M
96 Qwen3-VL-2B-Instruct Qwen · image-text-to-text 2.1 B ✓ apache-2.0 3.0 M from 5.5 GB
97 Florence-2-base microsoft · image-text-to-text 230 M ✓ mit 3.0 M from 1.0 GB
98 stable-diffusion-xl-base-1.0 stabilityai · text-to-image 2.6 B ⚠ openrail++ 3.0 M from 39.7 GB
99 all-MiniLM-L6-v2 Xenova · feature-extraction 20 M 512 ✓ apache-2.0 3.0 M from 0.5 GB
100 gemma-3-1b-it google · text-generation · gated 1.0 B ⚠ gemma 2.9 M from 2.8 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.