1,011 models · refreshed nightly

Models that run on RTX 4090 · 24 GB

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dOn RTX 4090 · 24 GB
301 Phi-4-multimodal-instruct microsoft · ASR 5.6 B 131 K ✓ mit 572 K 15.4 GB
302 gpt2-medium openai-community · text-generation 380 M ✓ mit 572 K 2.2 GB
303 wav2vec2-xls-r-parlaspeech-hr classla · ASR 320 M unknown 564 K 1.9 GB
304 gpt-oss-20b-GGUF unsloth · text-generation 131 K ✓ apache-2.0 560 K 13.1 GB
305 Hermes-3-Llama-3.1-8B NousResearch · text-generation 8.0 B 131 K ⚠ llama3 558 K 19.4 GB
306 sat-3l-sm segment-any-text · token-classification 210 M 514 ✓ mit 558 K 1.5 GB
307 msmarco-bert-base-dot-v5 sentence-transformers · sentence-similarity 110 M 512 unknown 557 K 1.0 GB
308 deepseek-coder-6.7b-instruct deepseek-ai · text-generation 6.7 B 16 K other 553 K 16.3 GB
309 Llama-3.2-3B meta-llama · text-generation · gated 3.2 B ⚠ llama3.2 552 K 8.0 GB
310 MiMo-7B-RL XiaomiMiMo · text-generation 7.8 B 33 K ✓ mit 552 K 18.9 GB
311 whisper-bemba-stt AbelZimba · ASR 240 M unknown 546 K 1.6 GB
312 Phi-3-mini-4k-instruct microsoft · text-generation 3.8 B 4 K ✓ mit 544 K 9.5 GB
313 detr-resnet-50 facebook · object-detection 40 M 1 K ✓ apache-2.0 544 K 0.7 GB
314 macbert4csc-base-chinese shibing624 · text-generation 100 M 512 ✓ apache-2.0 542 K 1.0 GB
315 gte-base thenlper · sentence-similarity 110 M 512 ✓ mit 541 K 0.8 GB
316 privacy-filter openai · token-classification 1.4 B 131 K ✓ apache-2.0 536 K 6.9 GB
317 Qwen2.5-Coder-7B-Instruct-GPTQ-Int4 Qwen · text-generation 7.6 B 33 K ✓ apache-2.0 536 K 7.8 GB
318 Qwen-AgentWorld-35B-A3B-GGUF unsloth · text-generation ✓ apache-2.0 534 K 13.1 GB
319 Qwen3-Embedding-4B-W4A16-G128 boboliu · feature-extraction 4.1 B 41 K ✓ apache-2.0 532 K 4.0 GB
320 jina-clip-v2 jinaai · feature-extraction 870 M ✗ cc-by-nc-4.0 530 K 2.5 GB
321 granite-speech-3.3-2b ibm-granite · ASR 3.0 B ✓ apache-2.0 529 K 13.9 GB
322 Qwen2.5-Coder-7B-Instruct-AWQ Qwen · text-generation 7.6 B 33 K ✓ apache-2.0 522 K 7.8 GB
323 inclusively-classification E-MIMIC · text-classification 110 M 512 ✗ cc-by-nc-sa-4.0 521 K 1.0 GB
324 whisper-medium-gguf handy-computer · ASR ✓ apache-2.0 518 K 1.1 GB
325 jina-embeddings-v5-text-nano jinaai · feature-extraction 210 M 8 K ✗ cc-by-nc-4.0 515 K 1.1 GB
326 wav2vec2-xls-r-300m-ftspeech saattrupdan · ASR 320 M other 512 K 1.9 GB
327 fullstop-punctuation-multilang-large oliverguhr · token-classification 560 M 512 ✓ mit 511 K 3.0 GB
328 vit_base_patch16_224.augreg2_in21k_ft_in1k timm · image-classification 90 M ✓ apache-2.0 510 K 0.9 GB
329 wide_resnet50_2.racm_in1k timm · image-classification 70 M ✓ apache-2.0 506 K 0.8 GB
330 gemma-2-9b-it-AWQ-INT4 hugging-quants · text-generation 9.2 B 8 K ⚠ gemma 504 K 8.7 GB
331 lambda unslothai · feature-extraction unknown 504 K 0.5 GB
332 t5gemma-s-s-prefixlm google · text-generation · gated 310 M ⚠ gemma 499 K 1.2 GB
333 resnet34.a1_in1k timm · image-classification 20 M ✓ apache-2.0 495 K 0.6 GB
334 bert-base-turkish-cased-mean-nli-stsb-tr emrecan · sentence-similarity 110 M 512 ✓ apache-2.0 495 K 1.0 GB
335 convnext_tiny.in12k_ft_in1k timm · image-classification 30 M ✓ apache-2.0 494 K 0.6 GB
336 SmolLM2-360M-Instruct HuggingFaceTB · text-generation 360 M 8 K ✓ apache-2.0 494 K 1.4 GB
337 llmlingua-2-bert-base-multilingual-cased-meetingbank microsoft · token-classification 180 M 512 ✓ apache-2.0 493 K 1.3 GB
338 vietnamese-bi-encoder bkai-foundation-models · sentence-similarity 130 M 256 ✓ apache-2.0 492 K 1.1 GB
339 tiny-mixtral TitanML · text-generation 250 M 131 K unknown 490 K 1.6 GB
340 Realistic_Vision_V5.1_noVAE SG161222 · text-to-image 860 M ⚠ creativeml-openrail-m 485 K 19.4 GB
341 Meta-Llama-3.1-8B-Instruct-FP8 RedHatAI · text-generation 8.0 B 131 K ⚠ llama3.1 479 K 11.7 GB
342 wav2vec2-large-xlsr-mvc-swahili eddiegulay · ASR 320 M ✓ apache-2.0 476 K 1.9 GB
343 bge-m3-spa-law-qa littlejohn-ai · sentence-similarity · gated 570 M ✓ apache-2.0 473 K 3.1 GB
344 typhoon2.5-qwen3-4b typhoon-ai · text-generation 4.0 B 262 K ✓ apache-2.0 468 K 10.0 GB
345 Qwen3-4B-AWQ Qwen · text-generation 4.0 B 41 K ✓ apache-2.0 463 K 4.0 GB
346 vram-16 unslothai · feature-extraction unknown 462 K 0.5 GB
347 pplx-embed-v1-0.6b perplexity-ai · feature-extraction 600 M 33 K ✓ mit 461 K 3.2 GB
348 bge-base-en BAAI · feature-extraction 110 M 512 ✓ mit 460 K 1.0 GB
349 e5-mistral-7b-instruct intfloat · feature-extraction 7.1 B 33 K ✓ mit 460 K 17.2 GB
350 Bonsai-27B-mlx-1bit prism-ml · text-generation 1.7 B ✓ apache-2.0 457 K 6.4 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.