1,011 models · refreshed nightly

Models that run on 2 × 24 GB

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dOn 2 × 24 GB
301 NVIDIA-Nemotron-3-Nano-4B-BF16 nvidia · text-generation 4.0 B 262 K other 739 K 9.8 GB
302 Qwen2.5-32B-Instruct-AWQ Qwen · text-generation 32.8 B 33 K ✓ apache-2.0 738 K 26.7 GB
303 bert-base-multilingual-uncased-sentiment nlptown · text-classification 170 M 512 ✓ mit 735 K 1.3 GB
304 Qwen2.5-Coder-1.5B-Instruct Qwen · text-generation 1.5 B 33 K ✓ apache-2.0 728 K 4.1 GB
305 Qwen3-4B-Base Qwen · text-generation 4.0 B 33 K ✓ apache-2.0 722 K 10.0 GB
306 sd-turbo stabilityai · text-to-image 870 M unknown 722 K 14.9 GB
307 BiRefNet ZhengPeng7 · image-segmentation 220 M ✓ mit 715 K 1.0 GB
308 RMBG-2.0 briaai · image-segmentation · gated 220 M other 702 K 1.5 GB
309 gemma-4-31B-it-NVFP4-turbo LilaRest · text-generation 32.5 B ✓ apache-2.0 702 K 26.6 GB
310 bert-large-cased-finetuned-conll03-english dbmdz · token-classification 330 M 512 unknown 699 K 2.0 GB
311 VibeVoice-ASR microsoft · ASR 8.7 B ✓ mit 695 K 20.9 GB
312 ruri-v3-310m cl-nagoya · sentence-similarity 310 M 8 K ✓ apache-2.0 695 K 1.9 GB
313 nb-wav2vec2-1b-nynorsk NbAiLab · ASR 960 M ✓ apache-2.0 693 K 4.9 GB
314 turn-detector livekit · text-classification 130 M 8 K other 690 K 1.1 GB
315 vit_tiny_patch16_224.augreg_in21k_ft_in1k timm · image-classification 10 M ✓ apache-2.0 689 K 0.5 GB
316 xlm-roberta-base-language-detection papluca · text-classification 280 M 512 ✓ mit 687 K 1.8 GB
317 Mistral-7B-v0.1 mistralai · text-generation 7.2 B 33 K ✓ apache-2.0 679 K 17.5 GB
318 vit_base_patch8_224.augreg2_in21k_ft_in1k timm · image-classification 90 M ✓ apache-2.0 679 K 0.9 GB
319 phi-4 microsoft · text-generation 14.7 B 16 K ✓ mit 671 K 34.9 GB
320 repvgg_a0.rvgg_in1k timm · image-classification 10 M ✓ mit 669 K 0.5 GB
321 vit-base-nsfw-detector AdamCodd · image-classification 90 M ✓ apache-2.0 660 K 0.9 GB
322 VibeVoice-Realtime-0.5B microsoft · text-to-speech 1.0 B ✓ mit 657 K 2.9 GB
323 Qwen3-32B-AWQ Qwen · text-generation 32.8 B 41 K ✓ apache-2.0 651 K 26.7 GB
324 bloom-560m bigscience · text-generation 560 M ⚠ bigscience-bloom-rail-1.0 628 K 1.8 GB
325 wav2vec2-large-xlsr-53-gender-recognition-librispeech alefiury · audio-classification 320 M ✓ apache-2.0 626 K 1.9 GB
326 mask2former-swin-large-ade-semantic facebook · image-segmentation 220 M other 625 K 1.5 GB
327 DeepSeek-R1-Distill-Qwen-1.5B deepseek-ai · text-generation 1.8 B 131 K ✓ mit 619 K 4.7 GB
328 OLMo-2-0425-1B allenai · text-generation 1.5 B 4 K ✓ apache-2.0 618 K 7.3 GB
329 NVIDIA-Nemotron-3-Nano-30B-A3B-FP8 nvidia · text-generation 31.6 B 262 K other 613 K 41.2 GB
330 parakeet-tdt-0.6b-v3-gguf handy-computer · ASR ✓ cc-by-4.0 606 K 1.0 GB
331 Qwen2-7B-Instruct Qwen · text-generation 7.6 B 33 K ✓ apache-2.0 606 K 18.4 GB
332 Qwen2.5-Coder-7B-Instruct-AWQ Orion-zhen · text-generation 7.6 B 33 K ✓ apache-2.0 604 K 7.9 GB
333 cryptobert ElKulako · text-classification 120 M 512 ✓ mit 602 K 1.1 GB
334 VLM2Vec-Full TIGER-Lab · text-generation 4.2 B 131 K ✓ apache-2.0 599 K 10.2 GB
335 deepseek-coder-7b-instruct-v1.5 deepseek-ai · text-generation 6.9 B 4 K other 599 K 16.7 GB
336 nsfw_image_detector Freepik · image-classification 90 M ✓ mit 591 K 0.7 GB
337 Qwen3-30B-A3B-Instruct-2507-FP8 Qwen · text-generation 30.5 B 262 K ✓ apache-2.0 588 K 39.4 GB
338 LFM2.5-1.2B-Instruct LiquidAI · text-generation 1.2 B 128 K other 584 K 3.3 GB
339 snowflake-arctic-embed-s Snowflake · sentence-similarity 30 M 512 ✓ apache-2.0 584 K 0.7 GB
340 Qwen3-Coder-30B-A3B-Instruct-AWQ-4bit cyankiwi · text-generation 5.3 B 262 K ✓ apache-2.0 583 K 21.2 GB
341 nb-wav2vec2-1b-bokmaal-v2 NbAiLab · ASR 960 M ✓ apache-2.0 581 K 4.9 GB
342 Qwen3-Reranker-4B-W4A16-G128 boboliu · text-classification 4.1 B 41 K ✓ apache-2.0 578 K 4.0 GB
343 encodec_24khz facebook · feature-extraction 20 M unknown 577 K 0.6 GB
344 gemma-2-2b-it google · text-generation · gated 2.6 B ⚠ gemma 576 K 6.6 GB
345 sentence-bert-base-ja-mean-tokens-v2 sonoisa · feature-extraction 110 M 512 cc-by-sa-4.0 575 K 1.0 GB
346 Qwen3-TTS-12Hz-1.7B-VoiceDesign Qwen · text-to-speech 1.9 B ✓ apache-2.0 574 K 5.8 GB
347 Phi-4-multimodal-instruct microsoft · ASR 5.6 B 131 K ✓ mit 572 K 15.4 GB
348 gpt2-medium openai-community · text-generation 380 M ✓ mit 572 K 2.2 GB
349 wav2vec2-xls-r-parlaspeech-hr classla · ASR 320 M unknown 564 K 1.9 GB
350 gpt-oss-20b-GGUF unsloth · text-generation 131 K ✓ apache-2.0 560 K 13.1 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.