Updated 2026-08-06 · ranked by real download data

Trending text-generation

Ranked nightly from our download snapshots of the Hugging Face catalog. Every entry shows its licence and what hardware it realistically needs.

#ModelParamsContextCommercial use30dMin VRAM
01 Qwen3-Coder-30B-A3B-Instruct-GGUF unsloth · text-generation ✓ apache-2.0 4.7 M
02 Laguna-S-2.1-NVFP4 poolside · text-generation 117.6 B 1.0 M openmdw-1.1 386 K from 276.8 GB
03 Qwen3Guard-Gen-4B Qwen · text-generation 4.4 B 33 K ✓ apache-2.0 440 K from 10.9 GB
04 GLM-5.2 zai-org · text-generation 753.3 B 1.0 M ✓ mit 2.2 M from 1,770.8 GB
05 Kimi-K2.7-Code-NVFP4 nvidia · text-generation other 836 K
06 gemma-2-9b-it-AWQ-INT4 hugging-quants · text-generation 9.2 B 8 K ⚠ gemma 504 K from 22.2 GB
07 llama-3.3-70b-instruct-awq casperhansen · text-generation 70.6 B 131 K ⚠ llama3.3 426 K from 166.3 GB
08 Qwen-72B Qwen · text-generation 72.3 B 33 K other 2.2 M from 170.4 GB
09 Olmo-3-7B-Instruct allenai · text-generation 7.3 B 66 K ✓ apache-2.0 434 K from 17.7 GB
10 Apertus-70B-Instruct-2509-quantized.w4a16 RedHatAI · text-generation 11.3 B 66 K ✓ apache-2.0 338 K from 27.1 GB
11 CodeLlama-7b-hf codellama · text-generation 6.7 B 16 K ⚠ llama2 356 K from 16.3 GB
12 MiniCPM5-1B openbmb · text-generation 1.1 B 131 K ✓ apache-2.0 926 K from 3.0 GB
13 Qwen3-Coder-Next-FP8-dynamic RedHatAI · text-generation 79.8 B 262 K ✓ apache-2.0 864 K from 187.9 GB
14 bloomz-560m bigscience · text-generation 560 M ⚠ bigscience-bloom-rail-1.0 1.2 M from 1.8 GB
15 Qwen3.6-35B-A3B-abliterated-v4 Bahushruth · text-generation 34.7 B 262 K ✓ apache-2.0 916 K from 82.0 GB
16 DeepSeek-R1-Distill-Qwen-14B deepseek-ai · text-generation 14.8 B 131 K ✓ mit 766 K from 35.2 GB
17 Qwen2.5-Coder-14B-Instruct-GPTQ-Int4 Qwen · text-generation 14.8 B 33 K ✓ apache-2.0 237 K from 35.2 GB
18 gemma-3-1b-it google · text-generation 1.0 B ⚠ gemma 4.7 M from 2.9 GB
19 Mistral-7B-Instruct-v0.2-AWQ TheBloke · text-generation 7.2 B 33 K ✓ apache-2.0 323 K from 17.5 GB
20 NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 nvidia · text-generation 560.5 B 262 K other 469 K from 1,317.7 GB
21 Qwen3-8B-AWQ Qwen · text-generation 8.2 B 41 K ✓ apache-2.0 1.1 M from 19.7 GB
22 Qwen3.6-27B-Text-NVFP4-MTP sakamakismile · text-generation 16.7 B ✓ apache-2.0 747 K from 39.7 GB
23 GLM-5.2-AWQ-INT4 cyankiwi · text-generation 753.3 B 1.0 M ✓ mit 378 K from 1,770.8 GB
24 deepseek-v4-gguf antirez · text-generation ✓ mit 875 K
25 t5gemma-s-s-prefixlm google · text-generation 310 M ⚠ gemma 499 K from 1.2 GB
Membership and ranking refresh nightly after the snapshot run. VRAM is an estimate for the smallest quantization at 8K context — methodology.