1,000 models · refreshed nightly

Text generation models

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dMin VRAM
351 Vikhr-Nemo-12B-Instruct-R-21-09-24 Vikhrmodels · text-generation 12.3 B 1.0 M ✓ apache-2.0 239 K from 30.8 GB
352 LLaDA-8B-Instruct GSAI-ML · text-generation 8.0 B ✓ mit 239 K from 19.3 GB
353 Mellum2-12B-A2.5B-Instruct-GGUF-Q8_0 JetBrains · text-generation ✓ apache-2.0 238 K from 14.7 GB
354 Ornith-1.0-397B ornith-ai · text-generation 396.8 B ✓ mit 236 K from 933.0 GB
355 Laguna-XS-2.1-GGUF poolside · text-generation openmdw-1.1 235 K from 22.8 GB
356 TinyStories-1M roneneldan · text-generation 1.0 M 2 K unknown 232 K from 0.5 GB
357 diffusiongemma-26B-A4B-it-NVFP4 nvidia · text-generation 14.4 B ✓ apache-2.0 232 K from 23.4 GB
358 Mistral-Small-24B-Instruct-2501-AWQ stelterlab · text-generation 23.6 B 33 K ✓ apache-2.0 232 K from 19.7 GB
359 LLaDA2.0-mini inclusionAI · text-generation 16.3 B 33 K ✓ apache-2.0 230 K from 38.7 GB
360 h2ovl-mississippi-2b h2oai · text-generation 2.2 B ✓ apache-2.0 230 K from 5.6 GB
361 Agents-A1-4B-Q8_0-GGUF InternScience · text-generation ✓ apache-2.0 230 K from 1.2 GB
362 Qwen3-0.6B-GGUF Qwen · text-generation 600 M 41 K ✓ apache-2.0 229 K from 1.3 GB
363 Mistral-7B-Instruct-v0.3-AWQ solidrust · text-generation 7.3 B 33 K ✓ apache-2.0 229 K from 6.2 GB
364 Yi-Coder-9B-Chat-GGUF MaziyarPanahi · text-generation unknown 227 K from 2.7 GB
365 Qwen2.5-14B-Instruct-GPTQ-Int4 Qwen · text-generation 14.8 B 33 K ✓ apache-2.0 227 K from 13.7 GB
366 Ornith-1.5-9B-MLX-6bit ornith-ai · text-generation 9.0 B unknown 226 K from 9.8 GB
367 Llama-3.2-3B-Instruct-GGUF unsloth · text-generation 131 K ⚠ llama3.2 224 K from 2.0 GB
368 Qwen2.5-1.5B-Instruct-GGUF Qwen · text-generation ✓ apache-2.0 224 K from 1.3 GB
369 MiniMax-M2.5-NVFP4 nvidia · text-generation 116.4 B 197 K other 222 K from 171.8 GB
370 hyenadna-medium-450k-seqlen-hf LongSafari · text-generation 28 M ✓ bsd-3-clause 221 K from 0.6 GB
371 GLM-4.7-Flash-AWQ-4bit cyankiwi · text-generation 32.1 B 203 K ✓ mit 221 K from 27.5 GB
372 Olmo-3-7B-Instruct-SFT allenai · text-generation 7.3 B 66 K ✓ apache-2.0 220 K from 17.7 GB
373 Llama-2-7b-hf NousResearch · text-generation 6.7 B 4 K unknown 220 K from 16.3 GB
374 automotive flywheel-ai · text-generation 34.7 B ✓ apache-2.0 219 K from 29.0 GB
375 Qwen3-14B-GGUF Qwen · text-generation ✓ apache-2.0 219 K from 10.4 GB
376 Kimi-K2-Instruct moonshotai · text-generation 1,026.5 B 131 K other 219 K from 1,286.6 GB
377 CAJAL-4B Agnuxo · text-generation ✓ apache-2.0 219 K from 3.5 GB
378 Llama-3.1-Nemotron-Nano-8B-v1 nvidia · text-generation 8.0 B 131 K other 218 K from 19.4 GB
379 granite-4.1-8b ibm-granite · text-generation 8.8 B 131 K ✓ apache-2.0 218 K from 21.2 GB
380 Qwen3.6-27B-NVFP4 nvidia · text-generation 18.2 B ✓ apache-2.0 217 K from 27.3 GB
381 GLM-4.7-Flash-GGUF unsloth · text-generation ✓ mit 216 K from 13.0 GB
382 Gemma-4-E4B-DECKARD-HERETIC-NVFP4 AEON-7 · text-generation 6.2 B ⚠ gemma 216 K from 12.6 GB
383 TinyLLama-v0 Maykeye · text-generation <0.1 M 2 K ✓ apache-2.0 215 K from 0.5 GB
384 DeepSeek-R1-Distill-Llama-8B deepseek-ai · text-generation 8.0 B 131 K ✓ mit 215 K from 19.4 GB
385 Qwen2-7B Qwen · text-generation 7.6 B 131 K ✓ apache-2.0 214 K from 18.4 GB
386 wildguard allenai · text-generation · gated 7.3 B ✓ apache-2.0 214 K from 17.5 GB
387 Phi-3-mini-128k-instruct microsoft · text-generation 3.8 B 131 K ✓ mit 214 K from 9.5 GB
388 NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 nvidia · text-generation 560.5 B 262 K other 213 K from 1,317.7 GB
389 GLM-4.7-Flash-MLX-8bit lmstudio-community · text-generation 29.9 B 203 K ✓ mit 213 K from 40.0 GB
390 Ornith-1.0-35B-AEON-Ultimate-Uncensored-BF16 AEON-7 · text-generation 35.1 B ✓ mit 211 K from 83.0 GB
391 Qwen3.5-35B-A3B-APEX-GGUF mudler · text-generation ✓ apache-2.0 208 K from 28.4 GB
392 Qwen3-30B-A3B-GPTQ-Int4 Qwen · text-generation 30.5 B 41 K ✓ apache-2.0 208 K from 23.7 GB
393 Qwen_Qwen3-Next-80B-A3B-Thinking-GGUF bartowski · text-generation ✓ apache-2.0 208 K from 18.7 GB
394 xlnet-base-cased xlnet · text-generation ✓ mit 208 K
395 DeepSeek-R1-Distill-Llama-70B-GGUF unsloth · text-generation 131 K ⚠ llama3.3 208 K from 29.5 GB
396 Bielik-11B-v3.0-Instruct-awq speakleash · text-generation 11.3 B 33 K ✓ apache-2.0 208 K from 9.0 GB
397 Qwen2.5-Math-1.5B Qwen · text-generation 1.5 B 4 K ✓ apache-2.0 208 K from 4.1 GB
398 Qwen3.6-35B-A3B-DFlash z-lab · text-generation 390 M 262 K ✓ apache-2.0 208 K from 1.4 GB
399 Qwen3Guard-Gen-4B Qwen · text-generation 4.4 B 33 K ✓ apache-2.0 207 K from 10.9 GB
400 Laguna-XS-2.1-NVFP4 poolside · text-generation 33.4 B 262 K openmdw-1.1 206 K from 29.3 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.