1,011 models · refreshed nightly

Models that run on 2 × 24 GB

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dOn 2 × 24 GB
51 bge-reranker-base BAAI · text-classification 280 M 512 ✓ mit 4.6 M 1.8 GB
52 Prompt-Guard-86M meta-llama · text-classification · gated 280 M ⚠ llama3.1 4.5 M 1.8 GB
53 Qwen3-ASR-0.6B Qwen · ASR 940 M ✓ apache-2.0 4.3 M 2.7 GB
54 granite-4.1-8b ibm-granite · text-generation 8.8 B 131 K ✓ apache-2.0 4.2 M 21.2 GB
55 Qwen3-VL-4B-Instruct Qwen · image-text-to-text 4.4 B ✓ apache-2.0 4.0 M 10.9 GB
56 granite-embedding-small-english-r2 ibm-granite · feature-extraction 50 M 8 K ✓ apache-2.0 3.8 M 0.6 GB
57 GLM-OCR zai-org · image-text-to-text 1.3 B ✓ mit 3.8 M 3.6 GB
58 Ornith-1.0-35B-GGUF deepreinforce-ai · text-generation ✓ mit 3.8 M 23.8 GB
59 Ornith-1.0-35B-GGUF ornith-ai · text-generation ✓ mit 3.8 M 23.8 GB
60 gemma-4-26B-A4B-it-AWQ-4bit cyankiwi · image-text-to-text 26.6 B ✓ apache-2.0 3.7 M 23.4 GB
61 distilbert-base-uncased-finetuned-sst-2-english distilbert · text-classification 70 M 512 ✓ apache-2.0 3.7 M 0.8 GB
62 all-MiniLM-L12-v2 sentence-transformers · sentence-similarity 30 M 512 ✓ apache-2.0 3.5 M 0.7 GB
63 Qwen3.6-27B-NVFP4 unsloth · image-text-to-text 21.2 B ✓ apache-2.0 3.4 M 29.4 GB
64 Qwen3-Embedding-4B Qwen · feature-extraction 4.0 B 41 K ✓ apache-2.0 3.4 M 10.0 GB
65 Qwen2.5-14B-Instruct Qwen · text-generation 14.8 B 33 K ✓ apache-2.0 3.3 M 35.2 GB
66 Qwen3-Embedding-8B Qwen · feature-extraction 7.6 B 41 K ✓ apache-2.0 3.3 M 18.3 GB
67 jina-embeddings-v3 jinaai · feature-extraction 570 M 8 K ✗ cc-by-nc-4.0 3.2 M 1.8 GB
68 Qwen3-4B-Instruct-2507 Qwen · text-generation 4.0 B 262 K ✓ apache-2.0 3.2 M 10.0 GB
69 Qwen3-VL-8B-Instruct-FP8 Qwen · image-text-to-text 8.8 B ✓ apache-2.0 3.2 M 13.5 GB
70 fairface_age_image_detection dima806 · image-classification 90 M ✓ apache-2.0 3.2 M 2.4 GB
71 Qwen2-1.5B-Instruct Qwen · text-generation 1.5 B 33 K ✓ apache-2.0 3.2 M 4.1 GB
72 chandra-ocr-2 datalab-to · image-text-to-text 5.3 B ⚠ openrail 3.2 M 12.9 GB
73 llava-1.5-7b-hf llava-hf · image-text-to-text 7.1 B ⚠ llama2 3.1 M 17.1 GB
74 indonesian-roberta-base-posp-tagger w11wo · token-classification 120 M 512 ✓ mit 3.1 M 1.1 GB
75 Qwen3.5-0.8B Qwen · image-text-to-text 870 M ✓ apache-2.0 3.1 M 2.6 GB
76 Qwen2-VL-2B-Instruct Qwen · image-text-to-text 2.2 B 33 K ✓ apache-2.0 3.0 M 5.7 GB
77 Qwen2.5-Coder-14B-Instruct Qwen · text-generation 14.8 B 33 K ✓ apache-2.0 3.0 M 35.2 GB
78 koelectra-small-v3-nsmc daekeun-ml · text-classification 10 M 512 ✓ mit 2.9 M 0.6 GB
79 Qwen3-14B Qwen · text-generation 14.8 B 41 K ✓ apache-2.0 2.8 M 35.2 GB
80 Florence-2-base microsoft · image-text-to-text 230 M ✓ mit 2.8 M 1.0 GB
81 resnet50.a1_in1k timm · image-classification 30 M ✓ apache-2.0 2.8 M 0.6 GB
82 Gemma-4-31B-IT-NVFP4 nvidia · text-generation 20.9 B other 2.8 M 39.5 GB
83 paraphrase-MiniLM-L6-v2 sentence-transformers · sentence-similarity 20 M 512 ✓ apache-2.0 2.8 M 0.6 GB
84 Unlimited-OCR baidu · image-text-to-text 3.3 B 33 K ✓ mit 2.7 M 8.3 GB
85 Qwen3.5-2B Qwen · image-text-to-text 2.3 B ✓ apache-2.0 2.6 M 5.8 GB
86 distilgpt2 distilbert · text-generation 90 M ✓ apache-2.0 2.6 M 0.9 GB
87 moondream2 vikhyatk · image-text-to-text 1.9 B ✓ apache-2.0 2.6 M 5.0 GB
88 Qwen3.6-27B-AWQ-INT4 cyankiwi · image-text-to-text 29.3 B ✓ apache-2.0 2.6 M 27.4 GB
89 paraphrase-MiniLM-L3-v2 sentence-transformers · sentence-similarity 20 M 512 ✓ apache-2.0 2.6 M 0.6 GB
90 Bonsai-27B-gguf prism-ml · text-generation ✓ apache-2.0 2.6 M 1.2 GB
91 Qwen2.5-0.5B Qwen · text-generation 490 M 33 K ✓ apache-2.0 2.6 M 1.7 GB
92 e5-large-v2 intfloat · sentence-similarity 340 M 512 ✓ mit 2.6 M 2.0 GB
93 TinyLlama-1.1B-Chat-v1.0 TinyLlama · text-generation 1.1 B 2 K ✓ apache-2.0 2.5 M 3.1 GB
94 snowflake-arctic-embed-xs Snowflake · sentence-similarity 20 M 512 unknown 2.5 M 0.6 GB
95 Qwen2.5-14B-Instruct-AWQ Qwen · text-generation 14.8 B 33 K ✓ apache-2.0 2.5 M 13.7 GB
96 chatterbox ResembleAI · text-to-speech ✓ mit 2.5 M 12.2 GB
97 mms-300m-1130-forced-aligner MahmoudAshraf · ASR 320 M ✗ cc-by-nc-4.0 2.5 M 1.9 GB
98 pythia-160m EleutherAI · text-generation 210 M 2 K ✓ apache-2.0 2.5 M 0.9 GB
99 bge-reranker-large BAAI · feature-extraction 560 M 512 ✓ mit 2.4 M 3.0 GB
100 Qwen3-14B-AWQ Qwen · text-generation 14.8 B 41 K ✓ apache-2.0 2.4 M 13.7 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.