1,000 models · refreshed nightly

All models

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dMin VRAM
251 bge-small-en-v1.5-onnx-Q Qdrant · sentence-similarity 512 ✓ apache-2.0 1.2 M
252 LFM2.5-2.6B-GGUF LiquidAI · text-generation other 1.2 M from 2.3 GB
253 surya-ocr-2 datalab-to · image-text-to-text 690 M ⚠ openrail 1.2 M from 2.1 GB
254 Ornith-1.5-35B-A3B-NVFP4 ornith-ai · text-generation 19.5 B ✓ mit 1.2 M from 29.2 GB
255 wavlm-large microsoft · feature-extraction unknown 1.2 M
256 Qwen3.8-27B-NVFP4 Inferact · image-text-to-text 17.6 B ✓ apache-2.0 1.2 M from 32.2 GB
257 PowerMoE-3b ibm-research · text-generation 3.4 B 4 K ✓ apache-2.0 1.2 M from 15.9 GB
258 tiny-gpt2 sshleifer · text-generation unknown 1.2 M
259 specter2_base allenai · feature-extraction 512 ✓ apache-2.0 1.2 M
260 whisper-tiny Xenova · ASR 40 M ✓ apache-2.0 1.2 M from 0.6 GB
261 pythia-70m-deduped EleutherAI · text-generation 100 M 2 K ✓ apache-2.0 1.2 M from 0.7 GB
262 Qwen3-VL-Embedding-8B Qwen · sentence-similarity 8.1 B ✓ apache-2.0 1.2 M from 19.6 GB
263 Qwen3.8-27B-GGUF ggml-org · image-text-to-text 27.0 B ✓ apache-2.0 1.2 M from 6.4 GB
264 Qwen3-VL-30B-A3B-Instruct-AWQ QuantTrio · text-generation 31.1 B ✓ apache-2.0 1.2 M from 24.8 GB
265 Ornith-1.0-9B ornith-ai · text-generation <0.1 M ✓ mit 1.2 M from 21.2 GB
266 Qwen2.5-Coder-14B-Instruct-AWQ Qwen · text-generation 14.8 B 33 K ✓ apache-2.0 1.1 M from 13.7 GB
267 Qwen3.6-27B-MTP-GGUF unsloth · image-text-to-text ✓ apache-2.0 1.1 M from 14.3 GB
268 finbert-tone yiyanghkust · text-classification 512 unknown 1.1 M
269 wav2vec2-xls-r-300m-bengali arijitx · ASR 300 M ✓ apache-2.0 1.1 M from 1.2 GB
270 Qwen3-VL-Embedding-2B Qwen · sentence-similarity 2.1 B ✓ apache-2.0 1.1 M from 5.5 GB
271 gemma-4-31B-it-NVFP4 RedHatAI · image-text-to-text 32.7 B ✓ apache-2.0 1.1 M from 31.0 GB
272 DeepSeek-V3 deepseek-ai · text-generation 684.5 B 164 K unknown 1.1 M from 860.6 GB
273 Qwen3-Coder-30B-A3B-Instruct-FP8 Qwen · text-generation 30.5 B 262 K ✓ apache-2.0 1.1 M from 39.4 GB
274 wav2vec2-xls-r-parlaspeech-hr classla · ASR 320 M unknown 1.1 M from 1.9 GB
275 all-MiniLM-L6-v2-onnx Qdrant · sentence-similarity 512 ✓ apache-2.0 1.1 M
276 Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V10-GGUF LuffyTheFox · image-text-to-text ✓ apache-2.0 1.1 M from 29.9 GB
277 medgemma-4b-it google · image-text-to-text · gated 4.3 B other 1.1 M from 10.6 GB
278 Qwen3.6-27B-GGUF unsloth · image-text-to-text ✓ apache-2.0 1.1 M from 14.1 GB
279 Qwen2.5-VL-32B-Instruct Qwen · image-text-to-text 33.5 B 128 K ✓ apache-2.0 1.1 M from 80.6 GB
280 OTel-LLM-27B-IT farbodtavakkoli · text-generation 27.0 B ✓ apache-2.0 1.1 M from 64.0 GB
281 clipseg-rd64-refined CIDAS · image-segmentation 150 M ✓ apache-2.0 1.1 M from 1.2 GB
282 surya-ocr-2-gguf datalab-to · image-text-to-text ⚠ openrail 1.1 M from 1.9 GB
283 UAE-Large-V1 WhereIsAI · feature-extraction 340 M 512 ✓ mit 1.1 M from 2.0 GB
284 robertuito-sentiment-analysis pysentimiento · text-classification 110 M 128 unknown 1.1 M from 1.0 GB
285 fairface_age_image_detection dima806 · image-classification 90 M ✓ apache-2.0 1.1 M from 2.4 GB
286 Muse-Glimmer-30B-GGUF unsloth · image-text-to-text ✓ apache-2.0 1.1 M from 12.3 GB
287 Bonsai-27B-mlx-1bit prism-ml · text-generation 1.7 B ✓ apache-2.0 1.1 M from 6.4 GB
288 PP-DocLayoutV3_safetensors PaddlePaddle · object-detection 30 M ✓ apache-2.0 1.1 M from 0.7 GB
289 gte-small thenlper · sentence-similarity 30 M 512 ✓ mit 1.1 M from 0.6 GB
290 Ternary-Bonsai-27B-mlx-2bit prism-ml · text-generation 2.6 B ✓ apache-2.0 1.1 M from 10.2 GB
291 parakeet-tdt-0.6b-v2 mlx-community · ASR 620 M ✓ cc-by-4.0 1.1 M from 3.3 GB
292 bge-large-zh-v1.5 BAAI · feature-extraction 512 ✓ mit 1.1 M
293 table-transformer-structure-recognition microsoft · object-detection 30 M 1 K ✓ mit 1.1 M from 0.6 GB
294 bm25 Qdrant · sentence-similarity ✓ apache-2.0 1.0 M
295 bge-micro-v2 TaylorAI · sentence-similarity 20 M 512 ✓ mit 1.0 M from 0.5 GB
296 Qwen3.6-35B-A3B-MTP-GGUF unsloth · image-text-to-text ✓ apache-2.0 1.0 M from 13.0 GB
297 Llama-3.1-8B-Instruct-4bit mlx-community · text-generation 1.3 B 131 K ⚠ llama3.1 1.0 M from 5.7 GB
298 1 unslothai · feature-extraction 2 K unknown 1.0 M from 0.5 GB
299 DeepSeek-V4-Flash-DSpark deepseek-ai · text-generation 165.3 B 1.0 M ✓ mit 1.0 M from 208.9 GB
300 NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4 nvidia · text-generation 17.8 B 1.0 M other 1.0 M from 26.9 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.