1,054 models · refreshed nightly

Text generation models

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dMin VRAM
51 Qwen3.8-27B-Heretic-Abliterated-Uncensored-GGUF 0bserverx · text-generation ✓ apache-2.0 1.8 M from 8.9 GB
52 Gemma-4-26B-A4B-NVFP4 nvidia · text-generation 14.4 B ✓ apache-2.0 1.8 M from 23.3 GB
53 Qwen3-30B-A3B Qwen · text-generation 30.5 B 41 K ✓ apache-2.0 1.8 M from 72.3 GB
54 Qwen3-8B-AWQ Qwen · text-generation 8.2 B 41 K ✓ apache-2.0 1.8 M from 8.4 GB
55 Mistral-7B-Instruct-v0.2 mistralai · text-generation 7.2 B 33 K ✓ apache-2.0 1.8 M from 17.5 GB
56 Gemma-4-31B-IT-NVFP4 nvidia · text-generation 20.9 B other 1.6 M from 39.5 GB
57 Qwen2.5-0.5B Qwen · text-generation 490 M 33 K ✓ apache-2.0 1.6 M from 1.7 GB
58 DeepSeek-V4-Flash deepseek-ai · text-generation 158.1 B 1.0 M ✓ mit 1.6 M from 199.8 GB
59 Qwen3-4B-Base Qwen · text-generation 4.0 B 33 K ✓ apache-2.0 1.5 M from 10.0 GB
60 Qwen3-1.7B-Base Qwen · text-generation 1.7 B 33 K ✓ apache-2.0 1.5 M from 4.5 GB
61 SmolLM2-135M-Instruct HuggingFaceTB · text-generation 130 M 8 K ✓ apache-2.0 1.5 M from 0.8 GB
62 Ternary-Bonsai-2-27B-gguf prism-ml · text-generation 27.0 B ✓ apache-2.0 1.5 M from 5.2 GB
63 OTel-LLM-E4B-IT farbodtavakkoli · text-generation 4.0 B ✓ apache-2.0 1.5 M from 9.9 GB
64 Qwen2.5-Coder-32B-Instruct-AWQ Qwen · text-generation 32.8 B 33 K ✓ apache-2.0 1.5 M from 26.7 GB
65 MiniMax-M2.7 MiniMaxAI · text-generation 228.7 B 205 K other 1.5 M from 288.0 GB
66 TinyLlama-1.1B-Chat-v1.0 TinyLlama · text-generation 1.1 B 2 K ✓ apache-2.0 1.5 M from 3.1 GB
67 Ornith-1.5-397B-GGUF ornith-ai · text-generation ✓ mit 1.4 M from 1.5 GB
68 Qwen3.5-122B-A10B-NVFP4 nvidia · text-generation 64.6 B ✓ apache-2.0 1.4 M from 102.0 GB
69 Qwen3-Coder-Next-FP8 Qwen · text-generation 79.7 B 262 K ✓ apache-2.0 1.4 M from 100.9 GB
70 OpenELM-1_1B-Instruct apple · text-generation 1.1 B apple-amlr 1.4 M from 3.0 GB
71 Qwen2.5-Coder-32B-Instruct Qwen · text-generation 32.8 B 33 K ✓ apache-2.0 1.3 M from 77.5 GB
72 PowerMoE-3b ibm-research · text-generation 3.4 B 4 K ✓ apache-2.0 1.3 M from 15.9 GB
73 NVIDIA-Nemotron-3-Super-120B-A12B-BF16 nvidia · text-generation 123.6 B 262 K other 1.3 M from 291.0 GB
74 gpt2-large openai-community · text-generation 810 M ✓ mit 1.3 M from 4.2 GB
75 Qwen3.8-27B-OBLITERATED OBLITERATUS · text-generation 27.8 B ✓ apache-2.0 1.3 M from 5.7 GB
76 Qwen2.5-Coder-14B-Instruct-AWQ Qwen · text-generation 14.8 B 33 K ✓ apache-2.0 1.2 M from 13.7 GB
77 LFM2.5-2.6B-GGUF LiquidAI · text-generation other 1.2 M from 2.3 GB
78 Meta-Llama-3-8B-Instruct meta-llama · text-generation · gated 8.0 B ⚠ llama3 1.2 M from 19.4 GB
79 Ornith-1.0-9B ornith-ai · text-generation <0.1 M ✓ mit 1.2 M from 21.2 GB
80 tiny-gpt2 sshleifer · text-generation unknown 1.2 M
81 Ornith-1.5-35B-A3B-NVFP4 ornith-ai · text-generation 19.5 B ✓ mit 1.2 M from 29.2 GB
82 Qwen3-VL-30B-A3B-Instruct-AWQ QuantTrio · text-generation 31.1 B ✓ apache-2.0 1.2 M from 24.8 GB
83 pythia-70m-deduped EleutherAI · text-generation 100 M 2 K ✓ apache-2.0 1.2 M from 0.7 GB
84 OTel-LLM-27B-IT farbodtavakkoli · text-generation 27.0 B ✓ apache-2.0 1.1 M from 64.0 GB
85 DeepSeek-V3 deepseek-ai · text-generation 684.5 B 164 K unknown 1.1 M from 860.6 GB
86 Qwen3-Coder-30B-A3B-Instruct-FP8 Qwen · text-generation 30.5 B 262 K ✓ apache-2.0 1.1 M from 39.4 GB
87 NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4 nvidia · text-generation 17.8 B 1.0 M other 1.1 M from 26.9 GB
88 Bonsai-27B-mlx-1bit prism-ml · text-generation 1.7 B ✓ apache-2.0 1.1 M from 6.4 GB
89 Ternary-Bonsai-27B-mlx-2bit prism-ml · text-generation 2.6 B ✓ apache-2.0 1.1 M from 10.2 GB
90 Llama-3.1-8B-Instruct-4bit mlx-community · text-generation 1.3 B 131 K ⚠ llama3.1 1.0 M from 5.7 GB
91 DeepSeek-V4-Flash-DSpark deepseek-ai · text-generation 165.3 B 1.0 M ✓ mit 1.0 M from 208.9 GB
92 glm-4-9b-chat-IMat-GGUF legraphista · text-generation other 1.0 M from 3.9 GB
93 Qwen2.5-32B-Instruct-AWQ Qwen · text-generation 32.8 B 33 K ✓ apache-2.0 1.0 M from 26.7 GB
94 DeepSeek-V3-0324 deepseek-ai · text-generation 684.5 B 164 K ✓ mit 991 K from 860.6 GB
95 JiRackUltra_14b CMSManhattan · text-generation 14.8 B 131 K ✓ mit 991 K from 9.1 GB
96 GLM-5.2-FP8 zai-org · text-generation 753.3 B 1.0 M ✓ mit 986 K from 944.7 GB
97 GLM-5.2 zai-org · text-generation 753.3 B 1.0 M ✓ mit 974 K from 1,770.8 GB
98 Qwen3-Coder-30B-A3B-Instruct-AWQ-4bit cyankiwi · text-generation 5.3 B 262 K ✓ apache-2.0 958 K from 21.2 GB
99 GLM-5.3 zai-org · text-generation 753.3 B 1.0 M other 947 K from 944.7 GB
100 Ornith-1.0-35B-FP8 deepreinforce-ai · text-generation 35.1 B ✓ mit 945 K from 47.2 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.