1,058 models · refreshed nightly

Models that run on RTX 4070 · 16 GB

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dOn RTX 4070 · 16 GB
351 DeepSeek-R1-Distill-Qwen-1.5B deepseek-ai · text-generation 1.8 B 131 K ✓ mit 431 K 4.7 GB
352 vit_small_patch16_224.augreg_in21k_ft_in1k timm · image-classification 20 M ✓ apache-2.0 430 K 0.6 GB
353 bert-base-turkish-cased-mean-nli-stsb-tr emrecan · sentence-similarity 110 M 512 ✓ apache-2.0 429 K 1.0 GB
354 whisper-large-v3-turbo-german primeline · ASR 810 M ✓ apache-2.0 428 K 2.4 GB
355 Qwen3-4B-GGUF Qwen · text-generation ✓ apache-2.0 428 K 3.2 GB
356 wav2vec2-large-mms-1b-azerbaijani-common_voice15.0 nijatzeynalov · ASR 960 M ✗ cc-by-nc-4.0 427 K 4.9 GB
357 Qwen2.5-3B-Instruct-AWQ Qwen · text-generation 3.4 B 33 K other 427 K 4.0 GB
358 llama-nemotron-embed-1b-v2 nvidia · feature-extraction 1.2 B 131 K other 425 K 3.4 GB
359 maple-preview-GGUF deepgrove · text-generation ✓ mit 424 K 6.5 GB
360 Qwen-AgentWorld-35B-A3B-GGUF unsloth · text-generation ✓ apache-2.0 423 K 13.1 GB
361 Bangla-twoclass-Sentiment-Analyzer Arunavaonly · text-classification 280 M 512 ✓ mit 422 K 1.8 GB
362 MiniCPM5-2B openbmb · text-generation 2.5 B 131 K ✓ apache-2.0 421 K 6.4 GB
363 Qwen3-4B-Thinking-2507 Qwen · text-generation 4.0 B 262 K ✓ apache-2.0 420 K 10.0 GB
364 wav2vec2-large-xls-r-300m-bg-d2 DrishtiSharma · ASR 300 M ✓ apache-2.0 419 K 1.2 GB
365 nemotron-speech-streaming-en-0.6b nvidia · ASR 618 M other 419 K 1.4 GB
366 Qwen3.6-14B-A3B-FableVibes-GGUF tvall43 · text-generation ✓ apache-2.0 419 K 6.4 GB
367 all-MiniLM-L6-v2-GGUF leliuga · sentence-similarity 20 M 512 ✓ apache-2.0 416 K 0.5 GB
368 vram-16 unslothai · feature-extraction unknown 410 K 0.5 GB
369 Phi-3-mini-4k-instruct microsoft · text-generation 3.8 B 4 K ✓ mit 407 K 9.5 GB
370 S-PubMedBert-MedQuAD TimKond · sentence-similarity 110 M 512 ✓ mit 406 K 1.0 GB
371 Qwen3-8B-GGUF Qwen · text-generation ✓ apache-2.0 404 K 6.0 GB
372 gemma-2-9b-it-AWQ-INT4 hugging-quants · text-generation 9.2 B 8 K ⚠ gemma 402 K 8.7 GB
373 wav2vec2-large-xlsr-marathi sumedh · ASR 320 M ✓ apache-2.0 399 K 1.9 GB
374 Qwen2.5-VL-7B-Instruct-NVFP4 nvidia · text-generation 5.0 B 128 K other 399 K 9.2 GB
375 jina-reranker-m0 jinaai · text-classification 2.4 B 33 K ✗ cc-by-nc-4.0 395 K 6.2 GB
376 gemma-3-270m google · text-generation · gated 270 M ⚠ gemma 395 K 1.1 GB
377 granite-4.1-3b ibm-granite · text-generation 3.4 B 131 K ✓ apache-2.0 393 K 8.5 GB
378 Ternary-Bonsai-8B-gguf prism-ml · text-generation ✓ apache-2.0 392 K 2.9 GB
379 Qwen3.8-27B-DFlash2 incoai · text-generation 1.9 B 262 K ✓ apache-2.0 391 K 5.0 GB
380 Kwaipilot_KAT-Coder-V2.5-Dev-GGUF bartowski · text-generation ✓ apache-2.0 390 K 11.3 GB
381 Phi-4-mini-instruct microsoft · text-generation 3.8 B 131 K ✓ mit 384 K 9.5 GB
382 TwIL-LM3 webAI-Official · text-generation 3.1 B 66 K other 384 K 3.1 GB
383 rtdetr_r50vd_coco_o365 PekingU · object-detection 40 M ✓ apache-2.0 380 K 0.7 GB
384 pythia-14m EleutherAI · text-generation 10 M 2 K ✓ apache-2.0 378 K 0.5 GB
385 mistral-7b-v0.3-bnb-4bit unsloth · text-generation 7.5 B 33 K ✓ apache-2.0 377 K 6.2 GB
386 Ornith-1.5-9B-OBLITERATED OBLITERATUS · text-generation 9.7 B ✓ mit 371 K 6.3 GB
387 segformer-b0-finetuned-ade-512-512 nvidia · image-segmentation <0.1 M other 368 K 0.5 GB
388 Llama-3.2-3B meta-llama · text-generation · gated 3.2 B ⚠ llama3.2 368 K 8.0 GB
389 Parable-Granite-4.1-3B-Claude-Fable-5-GGUF AnkitAI · text-generation ✓ apache-2.0 366 K 2.8 GB
390 privacy-filter-multilingual-GGUF LocalAI-io · token-classification ✓ apache-2.0 366 K 2.3 GB
391 Qwen3.8-27B-DFlash2 z-lab · text-generation 1.9 B 262 K ✓ apache-2.0 364 K 5.0 GB
392 Parable-Granite-4.1-8B-Claude-Fable-5-GGUF AnkitAI · text-generation ✓ apache-2.0 362 K 6.1 GB
393 Qwen2.5-Coder-1.5B Qwen · text-generation 1.5 B 33 K ✓ apache-2.0 361 K 4.1 GB
394 MiniCPM5-1B-Claude-Opus-Fable5-V2-Thinking-GGUF GnLOLot · text-generation ✓ apache-2.0 360 K 1.8 GB
395 POCKET-26B-GGUF FINAL-Bench · text-generation ✓ apache-2.0 360 K 12.7 GB
396 Z-Image-Turbo-GGUF unsloth · text-to-image ✓ apache-2.0 356 K 4.5 GB
397 vlt5-base-keywords Voicelab · text-generation 280 M ✓ cc-by-4.0 355 K 1.8 GB
398 t5-large google-t5 · translation 740 M ✓ apache-2.0 352 K 3.9 GB
399 amd.Instella-MoE-16B-A3B-Think-GGUF DevQuasar · text-generation unknown 351 K 7.7 GB
400 RMBG-1.4 briaai · image-segmentation 40 M other 349 K 0.7 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.