1,011 models · refreshed nightly

Models that run on RTX 3060 · 12 GB

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dOn RTX 3060 · 12 GB
251 Llama-3.2-3B meta-llama · text-generation · gated 3.2 B ⚠ llama3.2 552 K 8.0 GB
252 whisper-bemba-stt AbelZimba · ASR 240 M unknown 546 K 1.6 GB
253 Phi-3-mini-4k-instruct microsoft · text-generation 3.8 B 4 K ✓ mit 544 K 9.5 GB
254 detr-resnet-50 facebook · object-detection 40 M 1 K ✓ apache-2.0 544 K 0.7 GB
255 macbert4csc-base-chinese shibing624 · text-generation 100 M 512 ✓ apache-2.0 542 K 1.0 GB
256 gte-base thenlper · sentence-similarity 110 M 512 ✓ mit 541 K 0.8 GB
257 privacy-filter openai · token-classification 1.4 B 131 K ✓ apache-2.0 536 K 6.9 GB
258 Qwen2.5-Coder-7B-Instruct-GPTQ-Int4 Qwen · text-generation 7.6 B 33 K ✓ apache-2.0 536 K 7.8 GB
259 Qwen3-Embedding-4B-W4A16-G128 boboliu · feature-extraction 4.1 B 41 K ✓ apache-2.0 532 K 4.0 GB
260 jina-clip-v2 jinaai · feature-extraction 870 M ✗ cc-by-nc-4.0 530 K 2.5 GB
261 Qwen2.5-Coder-7B-Instruct-AWQ Qwen · text-generation 7.6 B 33 K ✓ apache-2.0 522 K 7.8 GB
262 inclusively-classification E-MIMIC · text-classification 110 M 512 ✗ cc-by-nc-sa-4.0 521 K 1.0 GB
263 whisper-medium-gguf handy-computer · ASR ✓ apache-2.0 518 K 1.1 GB
264 jina-embeddings-v5-text-nano jinaai · feature-extraction 210 M 8 K ✗ cc-by-nc-4.0 515 K 1.1 GB
265 wav2vec2-xls-r-300m-ftspeech saattrupdan · ASR 320 M other 512 K 1.9 GB
266 fullstop-punctuation-multilang-large oliverguhr · token-classification 560 M 512 ✓ mit 511 K 3.0 GB
267 vit_base_patch16_224.augreg2_in21k_ft_in1k timm · image-classification 90 M ✓ apache-2.0 510 K 0.9 GB
268 wide_resnet50_2.racm_in1k timm · image-classification 70 M ✓ apache-2.0 506 K 0.8 GB
269 gemma-2-9b-it-AWQ-INT4 hugging-quants · text-generation 9.2 B 8 K ⚠ gemma 504 K 8.7 GB
270 lambda unslothai · feature-extraction unknown 504 K 0.5 GB
271 t5gemma-s-s-prefixlm google · text-generation · gated 310 M ⚠ gemma 499 K 1.2 GB
272 resnet34.a1_in1k timm · image-classification 20 M ✓ apache-2.0 495 K 0.6 GB
273 bert-base-turkish-cased-mean-nli-stsb-tr emrecan · sentence-similarity 110 M 512 ✓ apache-2.0 495 K 1.0 GB
274 convnext_tiny.in12k_ft_in1k timm · image-classification 30 M ✓ apache-2.0 494 K 0.6 GB
275 SmolLM2-360M-Instruct HuggingFaceTB · text-generation 360 M 8 K ✓ apache-2.0 494 K 1.4 GB
276 llmlingua-2-bert-base-multilingual-cased-meetingbank microsoft · token-classification 180 M 512 ✓ apache-2.0 493 K 1.3 GB
277 vietnamese-bi-encoder bkai-foundation-models · sentence-similarity 130 M 256 ✓ apache-2.0 492 K 1.1 GB
278 tiny-mixtral TitanML · text-generation 250 M 131 K unknown 490 K 1.6 GB
279 Meta-Llama-3.1-8B-Instruct-FP8 RedHatAI · text-generation 8.0 B 131 K ⚠ llama3.1 479 K 11.7 GB
280 wav2vec2-large-xlsr-mvc-swahili eddiegulay · ASR 320 M ✓ apache-2.0 476 K 1.9 GB
281 bge-m3-spa-law-qa littlejohn-ai · sentence-similarity · gated 570 M ✓ apache-2.0 473 K 3.1 GB
282 typhoon2.5-qwen3-4b typhoon-ai · text-generation 4.0 B 262 K ✓ apache-2.0 468 K 10.0 GB
283 Qwen3-4B-AWQ Qwen · text-generation 4.0 B 41 K ✓ apache-2.0 463 K 4.0 GB
284 vram-16 unslothai · feature-extraction unknown 462 K 0.5 GB
285 pplx-embed-v1-0.6b perplexity-ai · feature-extraction 600 M 33 K ✓ mit 461 K 3.2 GB
286 bge-base-en BAAI · feature-extraction 110 M 512 ✓ mit 460 K 1.0 GB
287 Bonsai-27B-mlx-1bit prism-ml · text-generation 1.7 B ✓ apache-2.0 457 K 6.4 GB
288 Nemotron-3-Embed-1B-BF16 nvidia · sentence-similarity 1.1 B 262 K other 456 K 3.2 GB
289 Phi-4-mini-instruct microsoft · text-generation 3.8 B 131 K ✓ mit 455 K 9.5 GB
290 Bangla-twoclass-Sentiment-Analyzer Arunavaonly · text-classification 280 M 512 ✓ mit 454 K 1.8 GB
291 paraphrase-MiniLM-L12-v2 sentence-transformers · sentence-similarity 30 M 512 ✓ apache-2.0 454 K 0.7 GB
292 S-PubMedBert-MedQuAD TimKond · sentence-similarity 110 M 512 ✓ mit 453 K 1.0 GB
293 Ternary-Bonsai-27B-mlx-2bit prism-ml · text-generation 2.6 B ✓ apache-2.0 453 K 10.2 GB
294 Qwen2.5-3B-Instruct-unsloth-bnb-4bit unsloth · text-generation 3.2 B 33 K ✓ apache-2.0 449 K 3.6 GB
295 Qwen3-TTS-12Hz-0.6B-Base Qwen · text-to-speech 910 M ✓ apache-2.0 445 K 3.4 GB
296 Qwen3Guard-Gen-4B Qwen · text-generation 4.4 B 33 K ✓ apache-2.0 440 K 10.9 GB
297 vit_tiny_r_s16_p8_224.augreg_in21k timm · image-classification 10 M ✓ apache-2.0 432 K 0.5 GB
298 Qwen2.5-1.5B-apeach jason9693 · text-classification 1.5 B 131 K unknown 431 K 7.5 GB
299 Voxtral-Mini-4B-Realtime-2602-gguf handy-computer · ASR ✓ apache-2.0 430 K 3.6 GB
300 Zamba2-1.2B-instruct Zyphra · text-generation 1.2 B 4 K ✓ apache-2.0 429 K 6.0 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.