1,058 models · refreshed nightly

Models that run on RTX 4090 · 24 GB

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dOn RTX 4090 · 24 GB
351 Qwen2.5-Coder-7B-Instruct-GPTQ-Int4 Qwen · text-generation 7.6 B 33 K ✓ apache-2.0 515 K 7.8 GB
352 Realistic_Vision_V5.1_noVAE SG161222 · text-to-image 860 M ⚠ creativeml-openrail-m 509 K 19.4 GB
353 open-vakgyata onecxi · audio-classification 60 M ✗ cc-by-nc-4.0 507 K 0.8 GB
354 bloom-560m bigscience · text-generation 560 M ⚠ bigscience-bloom-rail-1.0 502 K 1.8 GB
355 wav2vec2-xls-r-300m-sk-cv8 comodoro · ASR 300 M ✓ apache-2.0 502 K 1.2 GB
356 lambda unslothai · feature-extraction unknown 498 K 0.5 GB
357 VibeVoice-1.5B microsoft · text-to-speech 2.7 B ✓ mit 497 K 6.9 GB
358 snowflake-arctic-embed-s Snowflake · sentence-similarity 30 M 512 ✓ apache-2.0 496 K 0.7 GB
359 pythia-160m-deduped EleutherAI · text-generation 210 M 2 K ✓ apache-2.0 492 K 0.9 GB
360 ruri-v3-310m cl-nagoya · sentence-similarity 310 M 8 K ✓ apache-2.0 492 K 1.9 GB
361 resnet34.a1_in1k timm · image-classification 20 M ✓ apache-2.0 492 K 0.6 GB
362 sat-3l-sm segment-any-text · token-classification 210 M 514 ✓ mit 491 K 1.5 GB
363 cryptobert ElKulako · text-classification 120 M 512 ✓ mit 483 K 1.1 GB
364 resnet18.a3_in1k timm · image-classification 10 M ✓ apache-2.0 483 K 0.6 GB
365 gemma-4-12B-coder-fable5-composer2.5-v1-GGUF yuxinlu1 · text-generation ✓ apache-2.0 483 K 5.8 GB
366 SmolLM2-360M HuggingFaceTB · text-generation 360 M 8 K ✓ apache-2.0 481 K 1.4 GB
367 snowflake-arctic-embed-l Snowflake · sentence-similarity 334 M 512 ✓ apache-2.0 480 K 2.0 GB
368 Parable-Qwen3-4B-Claude-Fable-5-GGUF AnkitAI · text-generation ✓ apache-2.0 479 K 3.2 GB
369 Qwen3-4B-GGUF unsloth · text-generation 41 K ✓ apache-2.0 474 K 2.3 GB
370 sentence-bert-base-ja-mean-tokens-v2 sonoisa · feature-extraction 110 M 512 cc-by-sa-4.0 474 K 1.0 GB
371 wav2vec2-large-xlsr-japanese-hiragana vumichien · ASR 320 M ✓ apache-2.0 472 K 1.9 GB
372 Apertus-8B-Instruct-2509 swiss-ai · text-generation 8.1 B 66 K ✓ apache-2.0 472 K 19.4 GB
373 wav2vec2-large-xls-r-300m-sinhala-low-LR-part1 SpideyDLK · ASR 320 M unknown 469 K 1.9 GB
374 wav2vec2-large-xlsr-53-basque stefan-it · ASR 320 M ✓ apache-2.0 468 K 1.9 GB
375 msmarco-distilbert-base-tas-b sentence-transformers · sentence-similarity 66 M 512 ✓ apache-2.0 463 K 0.8 GB
376 DeepSeek-Coder-V2-Lite-Instruct-GGUF bartowski · text-generation other 461 K 7.1 GB
377 pplx-embed-v1-0.6b perplexity-ai · feature-extraction 600 M 33 K ✓ mit 461 K 3.2 GB
378 mimi kyutai · feature-extraction 100 M 8 K ✓ cc-by-4.0 458 K 0.9 GB
379 MedCPT-Query-Encoder ncbi · feature-extraction 110 M 512 other 454 K 1.0 GB
380 vietnamese-bi-encoder bkai-foundation-models · sentence-similarity 130 M 256 ✓ apache-2.0 454 K 1.1 GB
381 indic-conformer-600m-multilingual ai4bharat · ASR · gated 600 M ✓ mit 453 K 1.9 GB
382 vit_base_patch16_224.augreg2_in21k_ft_in1k timm · image-classification 90 M ✓ apache-2.0 450 K 0.9 GB
383 gemma4-e4b-claims-comparison k-chirkunov · text-generation · gated 7.9 B ⚠ gemma 450 K 19.2 GB
384 Qwen3-8B-Base Qwen · text-generation 8.2 B 33 K ✓ apache-2.0 447 K 19.7 GB
385 Llama-2-7b-chat-hf meta-llama · text-generation · gated 6.7 B ⚠ llama2 447 K 16.3 GB
386 mask2former-swin-large-cityscapes-semantic facebook · image-segmentation 220 M other 444 K 1.5 GB
387 zeta-2.1-autoround-W4A16 LeaderboardModel1 · text-generation 2.2 B 33 K unknown 444 K 7.6 GB
388 msmarco-bert-base-dot-v5 sentence-transformers · sentence-similarity 110 M 512 unknown 444 K 1.0 GB
389 Voxtral-Mini-4B-Realtime-2602-gguf handy-computer · ASR ✓ apache-2.0 442 K 3.6 GB
390 Bonsai-27B-gguf prism-ml · text-generation ✓ apache-2.0 440 K 1.2 GB
391 BioLORD-2023 FremyCompany · sentence-similarity 110 M 512 other 437 K 1.0 GB
392 bge-small-en-v1.5 Xenova · feature-extraction 30 M 512 unknown 433 K 0.6 GB
393 DeepSeek-R1-Distill-Qwen-1.5B deepseek-ai · text-generation 1.8 B 131 K ✓ mit 431 K 4.7 GB
394 vit_small_patch16_224.augreg_in21k_ft_in1k timm · image-classification 20 M ✓ apache-2.0 430 K 0.6 GB
395 bert-base-turkish-cased-mean-nli-stsb-tr emrecan · sentence-similarity 110 M 512 ✓ apache-2.0 429 K 1.0 GB
396 whisper-large-v3-turbo-german primeline · ASR 810 M ✓ apache-2.0 428 K 2.4 GB
397 Qwen3-4B-GGUF Qwen · text-generation ✓ apache-2.0 428 K 3.2 GB
398 wav2vec2-large-mms-1b-azerbaijani-common_voice15.0 nijatzeynalov · ASR 960 M ✗ cc-by-nc-4.0 427 K 4.9 GB
399 Qwen2.5-3B-Instruct-AWQ Qwen · text-generation 3.4 B 33 K other 427 K 4.0 GB
400 llama-nemotron-embed-1b-v2 nvidia · feature-extraction 1.2 B 131 K other 425 K 3.4 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.