1,011 models · refreshed nightly

Models that run on RTX 4090 · 24 GB

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dOn RTX 4090 · 24 GB
201 e5-base intfloat · sentence-similarity 110 M 512 ✓ mit 990 K 1.0 GB
202 repeat unslothai · feature-extraction unknown 975 K 0.5 GB
203 roberta-base-go_emotions SamLowe · text-classification 120 M 512 ✓ mit 968 K 1.1 GB
204 Qwen3-1.7B-Base Qwen · text-generation 1.7 B 33 K ✓ apache-2.0 961 K 4.5 GB
205 Qwen3.6-27B-MLX-4bit lmstudio-community · image-text-to-text 4.7 B ✓ apache-2.0 942 K 18.9 GB
206 MiniCPM-V-4.6 openbmb · image-text-to-text 1.3 B ✓ apache-2.0 942 K 3.6 GB
207 robertuito-sentiment-analysis pysentimiento · text-classification 110 M 128 unknown 940 K 1.0 GB
208 Gemma-4-E4B-Uncensored-HauhauCS-Aggressive HauhauCS · image-text-to-text ⚠ gemma 937 K 5.4 GB
209 UAE-Large-V1 WhereIsAI · feature-extraction 340 M 512 ✓ mit 928 K 2.0 GB
210 MiniCPM5-1B openbmb · text-generation 1.1 B 131 K ✓ apache-2.0 926 K 3.0 GB
211 Qwen2.5-7B Qwen · text-generation 7.6 B 131 K ✓ apache-2.0 924 K 18.4 GB
212 LocateAnything-3B nvidia · image-text-to-text 3.8 B other 923 K 9.5 GB
213 rorshark-vit-base amunchet · image-classification 90 M ✓ apache-2.0 923 K 0.9 GB
214 Qwythos-9B-Claude-Mythos-5-1M-GGUF empero-ai · image-text-to-text ✓ apache-2.0 922 K 7.0 GB
215 InternVL2-1B OpenGVLab · image-text-to-text 940 M ✓ mit 920 K 2.7 GB
216 distilbert-base-multilingual-cased-sentiments-student lxyuan · text-classification 140 M 512 ✓ apache-2.0 917 K 1.1 GB
217 gemma-4-26B-A4B-it-QAT-MLX-4bit lmstudio-community · image-text-to-text 4.6 B ✓ apache-2.0 914 K 18.4 GB
218 Qwen2.5-1.5B-Instruct-AWQ Qwen · text-generation 1.8 B 33 K ✓ apache-2.0 913 K 2.5 GB
219 Phi-tiny-MoE-instruct microsoft · text-generation 3.8 B 4 K ✓ mit 907 K 9.3 GB
220 opus-mt-fr-en Helsinki-NLP · translation 80 M 512 ✓ apache-2.0 900 K 0.8 GB
221 VoxCPM2 openbmb · text-to-speech 2.3 B ✓ apache-2.0 900 K 5.9 GB
222 gte-large thenlper · sentence-similarity 340 M 512 ✓ mit 897 K 1.3 GB
223 gemma-4-26B-A4B-it-NVFP4 RedHatAI · image-text-to-text 15.1 B ✓ apache-2.0 897 K 20.8 GB
224 dreamshaper-7 Lykon · text-to-image 860 M ⚠ creativeml-openrail-m 895 K 9.7 GB
225 Qwen3.6-35B-A3B-GGUF unsloth · image-text-to-text ✓ apache-2.0 895 K 11.6 GB
226 Qwen3.6-27B-MLX-5bit lmstudio-community · image-text-to-text 5.5 B ✓ apache-2.0 892 K 22.7 GB
227 multi-qa-MiniLM-L6-cos-v1 sentence-transformers · sentence-similarity 20 M 512 unknown 882 K 0.6 GB
228 OmniVoice k2-fsa · text-to-speech 610 M unknown 877 K 4.2 GB
229 wav2vec2-large-xlsr-korean kresnik · ASR 320 M ✓ apache-2.0 877 K 1.9 GB
230 pubmedbert-base-embeddings NeuML · sentence-similarity 110 M 512 ✓ apache-2.0 877 K 1.0 GB
231 Qwen3-8B-FP8 Qwen · text-generation 8.2 B 41 K ✓ apache-2.0 876 K 12.1 GB
232 deepseek-v4-gguf antirez · text-generation ✓ mit 875 K 4.7 GB
233 SmolLM3-3B HuggingFaceTB · text-generation 3.1 B 66 K ✓ apache-2.0 872 K 7.7 GB
234 wav2vec2-large-xls-r-300m-Urdu kingabzpro · ASR 320 M ✓ apache-2.0 867 K 1.9 GB
235 Qwen2.5-1.5B Qwen · text-generation 1.5 B 131 K ✓ apache-2.0 863 K 4.1 GB
236 mamba-130m-hf state-spaces · text-generation 130 M unknown 846 K 1.1 GB
237 Qwen2.5-1.5B-quantized.w8a8 RedHatAI · text-generation 1.8 B 33 K ✓ apache-2.0 839 K 3.2 GB
238 llama-nemotron-embed-1b-v2 nvidia · feature-extraction 1.2 B 131 K other 828 K 3.4 GB
239 text2vec-base-chinese shibing624 · sentence-similarity 100 M 512 ✓ apache-2.0 824 K 1.0 GB
240 gte-small thenlper · sentence-similarity 30 M 512 ✓ mit 823 K 0.6 GB
241 PP-DocLayoutV3_safetensors PaddlePaddle · object-detection 30 M ✓ apache-2.0 822 K 0.7 GB
242 Llama-3.2-1B-Instruct-FP8 RedHatAI · text-generation 1.5 B 131 K ⚠ llama3.2 799 K 3.0 GB
243 Kimi-K3-DSpark RadixArk · text-generation 2.3 B 1.0 M unknown 798 K 5.8 GB
244 mimi kyutai · feature-extraction 100 M 8 K ✓ cc-by-4.0 798 K 0.9 GB
245 Flux2-Klein-9B-True-V2 wikeeyang · text-to-image other 792 K 6.8 GB
246 Llama-2-7b-hf meta-llama · text-generation · gated 6.7 B ⚠ llama2 779 K 16.3 GB
247 vntl-llama3-8b-v2-gguf lmg-anon · translation ⚠ llama3 774 K 6.8 GB
248 phi-2 microsoft · text-generation 2.8 B 2 K ✓ mit 771 K 7.0 GB
249 F5-TTS SWivid · text-to-speech ✗ cc-by-nc-4.0 770 K 5.0 GB
250 pythia-70m-deduped EleutherAI · text-generation 100 M 2 K ✓ apache-2.0 769 K 0.7 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.