1,011 models · refreshed nightly

Models that run on 2 × 24 GB

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dOn 2 × 24 GB
351 Hermes-3-Llama-3.1-8B NousResearch · text-generation 8.0 B 131 K ⚠ llama3 558 K 19.4 GB
352 sat-3l-sm segment-any-text · token-classification 210 M 514 ✓ mit 558 K 1.5 GB
353 DeepSeek-Coder-V2-Lite-Instruct deepseek-ai · text-generation 15.7 B 164 K other 557 K 37.4 GB
354 msmarco-bert-base-dot-v5 sentence-transformers · sentence-similarity 110 M 512 unknown 557 K 1.0 GB
355 deepseek-coder-6.7b-instruct deepseek-ai · text-generation 6.7 B 16 K other 553 K 16.3 GB
356 Llama-3.2-3B meta-llama · text-generation · gated 3.2 B ⚠ llama3.2 552 K 8.0 GB
357 MiMo-7B-RL XiaomiMiMo · text-generation 7.8 B 33 K ✓ mit 552 K 18.9 GB
358 whisper-bemba-stt AbelZimba · ASR 240 M unknown 546 K 1.6 GB
359 Phi-3-mini-4k-instruct microsoft · text-generation 3.8 B 4 K ✓ mit 544 K 9.5 GB
360 detr-resnet-50 facebook · object-detection 40 M 1 K ✓ apache-2.0 544 K 0.7 GB
361 macbert4csc-base-chinese shibing624 · text-generation 100 M 512 ✓ apache-2.0 542 K 1.0 GB
362 gte-base thenlper · sentence-similarity 110 M 512 ✓ mit 541 K 0.8 GB
363 privacy-filter openai · token-classification 1.4 B 131 K ✓ apache-2.0 536 K 6.9 GB
364 Qwen2.5-Coder-7B-Instruct-GPTQ-Int4 Qwen · text-generation 7.6 B 33 K ✓ apache-2.0 536 K 7.8 GB
365 Qwen-AgentWorld-35B-A3B-GGUF unsloth · text-generation ✓ apache-2.0 534 K 13.1 GB
366 Qwen3-Embedding-4B-W4A16-G128 boboliu · feature-extraction 4.1 B 41 K ✓ apache-2.0 532 K 4.0 GB
367 jina-clip-v2 jinaai · feature-extraction 870 M ✗ cc-by-nc-4.0 530 K 2.5 GB
368 granite-speech-3.3-2b ibm-granite · ASR 3.0 B ✓ apache-2.0 529 K 13.9 GB
369 Qwen2.5-Coder-7B-Instruct-AWQ Qwen · text-generation 7.6 B 33 K ✓ apache-2.0 522 K 7.8 GB
370 inclusively-classification E-MIMIC · text-classification 110 M 512 ✗ cc-by-nc-sa-4.0 521 K 1.0 GB
371 EXAONE-3.5-7.8B-Instruct LGAI-EXAONE · text-generation 7.8 B 33 K other 518 K 36.1 GB
372 whisper-medium-gguf handy-computer · ASR ✓ apache-2.0 518 K 1.1 GB
373 jina-embeddings-v5-text-nano jinaai · feature-extraction 210 M 8 K ✗ cc-by-nc-4.0 515 K 1.1 GB
374 wav2vec2-xls-r-300m-ftspeech saattrupdan · ASR 320 M other 512 K 1.9 GB
375 fullstop-punctuation-multilang-large oliverguhr · token-classification 560 M 512 ✓ mit 511 K 3.0 GB
376 vit_base_patch16_224.augreg2_in21k_ft_in1k timm · image-classification 90 M ✓ apache-2.0 510 K 0.9 GB
377 wide_resnet50_2.racm_in1k timm · image-classification 70 M ✓ apache-2.0 506 K 0.8 GB
378 gemma-2-9b-it-AWQ-INT4 hugging-quants · text-generation 9.2 B 8 K ⚠ gemma 504 K 8.7 GB
379 lambda unslothai · feature-extraction unknown 504 K 0.5 GB
380 DeepSeek-V2-Lite deepseek-ai · text-generation 15.7 B 164 K other 503 K 37.4 GB
381 t5gemma-s-s-prefixlm google · text-generation · gated 310 M ⚠ gemma 499 K 1.2 GB
382 resnet34.a1_in1k timm · image-classification 20 M ✓ apache-2.0 495 K 0.6 GB
383 bert-base-turkish-cased-mean-nli-stsb-tr emrecan · sentence-similarity 110 M 512 ✓ apache-2.0 495 K 1.0 GB
384 convnext_tiny.in12k_ft_in1k timm · image-classification 30 M ✓ apache-2.0 494 K 0.6 GB
385 SmolLM2-360M-Instruct HuggingFaceTB · text-generation 360 M 8 K ✓ apache-2.0 494 K 1.4 GB
386 llmlingua-2-bert-base-multilingual-cased-meetingbank microsoft · token-classification 180 M 512 ✓ apache-2.0 493 K 1.3 GB
387 Qwen2.5-32B-Instruct-GPTQ-Int4 Qwen · text-generation 32.8 B 33 K ✓ apache-2.0 492 K 26.7 GB
388 vietnamese-bi-encoder bkai-foundation-models · sentence-similarity 130 M 256 ✓ apache-2.0 492 K 1.1 GB
389 tiny-mixtral TitanML · text-generation 250 M 131 K unknown 490 K 1.6 GB
390 GLM-4.7-Flash-AWQ-4bit cyankiwi · text-generation 32.1 B 203 K ✓ mit 489 K 27.5 GB
391 Realistic_Vision_V5.1_noVAE SG161222 · text-to-image 860 M ⚠ creativeml-openrail-m 485 K 19.4 GB
392 Meta-Llama-3.1-8B-Instruct-FP8 RedHatAI · text-generation 8.0 B 131 K ⚠ llama3.1 479 K 11.7 GB
393 wav2vec2-large-xlsr-mvc-swahili eddiegulay · ASR 320 M ✓ apache-2.0 476 K 1.9 GB
394 bge-m3-spa-law-qa littlejohn-ai · sentence-similarity · gated 570 M ✓ apache-2.0 473 K 3.1 GB
395 typhoon2.5-qwen3-4b typhoon-ai · text-generation 4.0 B 262 K ✓ apache-2.0 468 K 10.0 GB
396 Bielik-11B-v3.0-Instruct speakleash · text-generation · gated 11.2 B ✓ apache-2.0 464 K 26.7 GB
397 Qwen3-4B-AWQ Qwen · text-generation 4.0 B 41 K ✓ apache-2.0 463 K 4.0 GB
398 vram-16 unslothai · feature-extraction unknown 462 K 0.5 GB
399 pplx-embed-v1-0.6b perplexity-ai · feature-extraction 600 M 33 K ✓ mit 461 K 3.2 GB
400 bge-base-en BAAI · feature-extraction 110 M 512 ✓ mit 460 K 1.0 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.