1,000 models · refreshed nightly

Models that run on RTX 4070 · 16 GB

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dOn RTX 4070 · 16 GB
351 mms-lid-126 facebook · audio-classification 970 M ✗ cc-by-nc-4.0 291 K 4.9 GB
352 deberta-v3-base-prompt-injection-v2 protectai · text-classification 180 M 512 ✓ apache-2.0 291 K 1.3 GB
353 CommunityForensics-DeepfakeDet-ViT buildborderless · image-classification 40 M ✓ mit 290 K 0.7 GB
354 rtdetr_v2_r18vd PekingU · object-detection 20 M ✓ apache-2.0 289 K 0.6 GB
355 SmolLM-135M HuggingFaceTB · text-generation 130 M 2 K ✓ apache-2.0 288 K 1.1 GB
356 Llama-3.2-1B-Instruct unsloth · text-generation 1.2 B 131 K ⚠ llama3.2 287 K 3.4 GB
357 MuQ-large-msd-iter OpenMuQ · audio-classification 330 M ✗ cc-by-nc-4.0 286 K 2.0 GB
358 gemma-4-12B-coder-fable5-composer2.5-v1-GGUF yuxinlu1 · text-generation ✓ apache-2.0 286 K 5.8 GB
359 roberta-large-mnli FacebookAI · text-classification 360 M 512 ✓ mit 285 K 2.1 GB
360 vit_small_patch16_224.augreg_in21k_ft_in1k timm · image-classification 20 M ✓ apache-2.0 283 K 0.6 GB
361 Qwen2.5-7B-Instruct-GPTQ-Int4 Qwen · text-generation 7.6 B 33 K ✓ apache-2.0 282 K 7.8 GB
362 nsfw-classifier giacomoarienti · image-classification 90 M ✗ cc-by-nc-nd-4.0 279 K 0.9 GB
363 DeepSeek-R1-0528-Qwen3-8B-MLX-4bit lmstudio-community · text-generation 1.3 B 131 K ✓ mit 278 K 5.8 GB
364 Qwen3-8B-GGUF unsloth · text-generation 41 K ✓ apache-2.0 277 K 4.1 GB
365 open-vakgyata onecxi · audio-classification 60 M ✗ cc-by-nc-4.0 274 K 0.8 GB
366 Qwen3-4B-GGUF Qwen · text-generation ✓ apache-2.0 274 K 3.2 GB
367 MiniCPM5-1B-Claude-Opus-Fable5-V2-Thinking-GGUF GnLOLot · text-generation ✓ apache-2.0 265 K 1.8 GB
368 Qwen2.5-Math-1.5B Qwen · text-generation 1.5 B 4 K ✓ apache-2.0 265 K 4.1 GB
369 madlad400-3b-mt google · translation 2.9 B ✓ apache-2.0 263 K 2.8 GB
370 gemma-2-2b google · text-generation · gated 2.6 B ⚠ gemma 263 K 12.4 GB
371 wikineural-multilingual-ner Babelscape · token-classification 180 M 512 ✗ cc-by-nc-sa-4.0 263 K 1.3 GB
372 DeepSeek-R1-0528-Qwen3-8B-MLX-8bit lmstudio-community · text-generation 2.3 B 131 K ✓ mit 258 K 10.4 GB
373 plant-identity umutbozdag · image-classification 90 M unknown 257 K 0.9 GB
374 t5-3b google-t5 · translation 2.9 B ✓ apache-2.0 255 K 13.5 GB
375 Sugoi-32B-Ultra-GGUF sugoitoolkit · translation ✓ apache-2.0 254 K 14.0 GB
376 bert-base-multilingual-cased-ner-hrl Davlan · token-classification 180 M 512 ✓ afl-3.0 249 K 1.3 GB
377 span-marker-bert-base-uncased-acronyms tomaarsen · token-classification 110 M ✓ apache-2.0 247 K 1.0 GB
378 Bielik-11B-v3.0-Instruct-awq speakleash · text-generation 11.3 B 33 K ✓ apache-2.0 247 K 9.0 GB
379 xlm-roberta-large-ner-hrl Davlan · token-classification 560 M 512 ✓ afl-3.0 246 K 3.0 GB
380 wav2vec2-large-robust-24-ft-age-gender audeering · audio-classification 320 M ✗ cc-by-nc-sa-4.0 246 K 1.9 GB
381 pythia-410m EleutherAI · text-generation 510 M 2 K ✓ apache-2.0 244 K 1.6 GB
382 llama-160m JackFram · text-generation 160 M 2 K ✓ apache-2.0 243 K 1.2 GB
383 hubert-large-speech-emotion-recognition-russian-dusha-finetuned xbgoose · audio-classification 320 M ✓ apache-2.0 243 K 1.9 GB
384 Phi-3-mini-128k-instruct microsoft · text-generation 3.8 B 131 K ✓ mit 242 K 9.5 GB
385 ai-image-detector-dev-deploy haywoodsloan · image-classification 200 M unknown 241 K 2.2 GB
386 table-transformer-structure-recognition-v1.1-all microsoft · object-detection 30 M ✓ mit 240 K 0.6 GB
387 Qwen3-8B-GGUF Qwen · text-generation ✓ apache-2.0 239 K 6.0 GB
388 tiny_starcoder_py bigcode · text-generation 160 M bigcode-openrail-m 238 K 1.2 GB
389 Qwen2.5-3B-Instruct-AWQ Qwen · text-generation 3.4 B 33 K other 238 K 4.0 GB
390 pythia-14m EleutherAI · text-generation 10 M 2 K ✓ apache-2.0 237 K 0.5 GB
391 internlm2-1_8b-reward internlm · text-classification 1.7 B 33 K other 236 K 4.5 GB
392 convnext_base.fb_in22k_ft_in1k timm · image-classification 90 M ✓ apache-2.0 236 K 0.9 GB
393 Qwen2.5-Coder-14B-Instruct-GPTQ-Int4 Qwen · text-generation 14.8 B 33 K ✓ apache-2.0 234 K 13.7 GB
394 layoutreader hantian · token-classification 360 M 514 ✗ cc-by-nc-sa-4.0 234 K 1.3 GB
395 gliner_multi-v2.1 urchade · token-classification 290 M ✓ apache-2.0 234 K 1.8 GB
396 Llama-3.2-1B-Instruct-Q8_0-GGUF hugging-quants · text-generation unknown 233 K 2.0 GB
397 distilbert-NER dslim · token-classification 70 M 512 ✓ apache-2.0 231 K 0.8 GB
398 Phi-3-vision-128k-instruct microsoft · text-generation 4.2 B 131 K ✓ mit 231 K 10.2 GB
399 TinyLLama-v0 Maykeye · text-generation 0 M 2 K ✓ apache-2.0 226 K 0.5 GB
400 LFM2.5-1.2B-Instruct-GGUF LiquidAI · text-generation other 226 K 1.3 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.