1,000 models · refreshed nightly

Text generation models

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dMin VRAM
101 Qwen2.5-Coder-7B-Instruct-AWQ Qwen · text-generation 7.6 B 33 K ✓ apache-2.0 930 K from 7.8 GB
102 Qwen3-4B-Instruct-2507-FP8 Qwen · text-generation 4.4 B 262 K ✓ apache-2.0 930 K from 6.9 GB
103 Llama-3.3-70B-Instruct meta-llama · text-generation · gated 70.6 B ⚠ llama3.3 913 K from 166.3 GB
104 DeepSeek-Coder-V2-Lite-Instruct deepseek-ai · text-generation 15.7 B 164 K other 911 K from 37.4 GB
105 Qwen3.6-35B-A3B-abliterated-v4 Bahushruth · text-generation 34.7 B 262 K ✓ apache-2.0 877 K from 82.0 GB
106 Llama-3.2-1B meta-llama · text-generation · gated 1.2 B ⚠ llama3.2 869 K from 3.4 GB
107 Qwen2.5-1.5B Qwen · text-generation 1.5 B 131 K ✓ apache-2.0 860 K from 4.1 GB
108 DeepSeek-R1-0528-Qwen3-8B deepseek-ai · text-generation 8.2 B 131 K ✓ mit 858 K from 19.7 GB
109 NVIDIA-Nemotron-3-Nano-30B-A3B-NVFP4 nvidia · text-generation 18.2 B 262 K other 827 K from 24.5 GB
110 Qwen3-0.6B-Base Qwen · text-generation 600 M 33 K ✓ apache-2.0 816 K from 1.9 GB
111 POCKET-35B-GGUF FINAL-Bench · text-generation ✓ apache-2.0 812 K from 9.6 GB
112 Qwen3-32B-AWQ Qwen · text-generation 32.8 B 41 K ✓ apache-2.0 811 K from 26.7 GB
113 Llama-2-7b-hf meta-llama · text-generation · gated 6.7 B ⚠ llama2 806 K from 16.3 GB
114 Ornith-1.5-9B-NVFP4 ornith-ai · text-generation 6.7 B ✓ mit 805 K from 11.2 GB
115 Qwen3-Coder-480B-A35B-Instruct-FP8 Qwen · text-generation 480.2 B 262 K ✓ apache-2.0 804 K from 602.9 GB
116 OLMo-2-0425-1B allenai · text-generation 1.5 B 4 K ✓ apache-2.0 799 K from 7.3 GB
117 Qwen3.8-4B-Distill-GGUF empero-ai · text-generation ✓ apache-2.0 793 K from 3.6 GB
118 Qwen3-8B-FP8 Qwen · text-generation 8.2 B 41 K ✓ apache-2.0 780 K from 12.1 GB
119 DeepSeek-R1 deepseek-ai · text-generation 684.5 B 164 K ✓ mit 773 K from 860.6 GB
120 Qwen3.8-2B-Distill-GGUF empero-ai · text-generation ✓ apache-2.0 772 K from 1.9 GB
121 Qwen3-30B-A3B-Instruct-2507 Qwen · text-generation 30.5 B 262 K ✓ apache-2.0 768 K from 72.3 GB
122 pythia-6.9b EleutherAI · text-generation 7.0 B 2 K ✓ apache-2.0 754 K from 16.8 GB
123 GLM-5.2-NVFP4 nvidia · text-generation 381.0 B 1.0 M ✓ mit 751 K from 569.0 GB
124 gemma-2-9b-it google · text-generation · gated 9.2 B ⚠ gemma 745 K from 22.2 GB
125 NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4 nvidia · text-generation 67.2 B 262 K other 744 K from 98.9 GB
126 bloomz-560m bigscience · text-generation 560 M ⚠ bigscience-bloom-rail-1.0 724 K from 1.8 GB
127 deepseek-coder-7b-instruct-v1.5 deepseek-ai · text-generation 6.9 B 4 K other 718 K from 16.7 GB
128 Qwen3-14B-NVFP4 nvidia · text-generation 8.2 B 41 K ✓ apache-2.0 713 K from 13.3 GB
129 Qwen2-0.5B Qwen · text-generation 490 M 131 K ✓ apache-2.0 709 K from 1.7 GB
130 gemma-4-12B-agentic-fable5-composer2.5-v2-3.5x-tau2-GGUF yuxinlu1 · text-generation ✓ apache-2.0 708 K from 1.4 GB
131 Qwen2.5-Coder-7B Qwen · text-generation 7.6 B 33 K ✓ apache-2.0 706 K from 18.4 GB
132 granite-4.1-30b ibm-granite · text-generation 28.9 B 131 K ✓ apache-2.0 700 K from 68.3 GB
133 Qwen3.8-27B-DFlash2-GGUF z-lab · text-generation ✓ apache-2.0 696 K from 1.8 GB
134 Qwen3.8-9B-Distill-GGUF empero-ai · text-generation ✓ apache-2.0 693 K from 6.9 GB
135 Qwen2.5-7B Qwen · text-generation 7.6 B 131 K ✓ apache-2.0 687 K from 18.4 GB
136 Qwen1.5-MoE-A2.7B Qwen · text-generation 14.3 B 8 K other 673 K from 34.1 GB
137 Ornith-1.0-397B-FP8 deepreinforce-ai · text-generation 397.0 B ✓ mit 673 K from 505.7 GB
138 NVIDIA-Nemotron-3-Nano-30B-A3B-BF16 nvidia · text-generation 31.6 B 262 K other 666 K from 74.7 GB
139 gemma-2-2b-it google · text-generation · gated 2.6 B ⚠ gemma 663 K from 6.6 GB
140 Ternary-Bonsai-27B-gguf prism-ml · text-generation ✓ apache-2.0 663 K from 1.2 GB
141 GLM-5.3-Flash-GGUF unsloth · text-generation ✓ mit 661 K from 54.7 GB
142 gpt-neox-20b EleutherAI · text-generation 20.7 B 2 K ✓ apache-2.0 652 K from 49.0 GB
143 phi-4 microsoft · text-generation 14.7 B 16 K ✓ mit 651 K from 34.9 GB
144 Ornith-1.5-397B ornith-ai · text-generation 403.4 B ✓ mit 636 K from 948.5 GB
145 Qwen2.5-Coder-7B-Instruct-AWQ Orion-zhen · text-generation 7.6 B 33 K ✓ apache-2.0 623 K from 7.9 GB
146 MiniCPM-SALA-AWQ-8bit cyankiwi · text-generation 3.1 B 524 K ✓ apache-2.0 616 K from 12.7 GB
147 SmolLM3-3B-Base HuggingFaceTB · text-generation 3.1 B 66 K ✓ apache-2.0 615 K from 7.7 GB
148 SmolLM3-3B HuggingFaceTB · text-generation 3.1 B 66 K ✓ apache-2.0 614 K from 7.7 GB
149 EXAONE-3.5-7.8B-Instruct-AWQ LGAI-EXAONE · text-generation 7.8 B 33 K other 614 K from 7.5 GB
150 MiniCPM5-1B openbmb · text-generation 1.1 B 131 K ✓ apache-2.0 611 K from 3.0 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.