1,011 models · refreshed nightly

Models that run on RTX 4090 · 24 GB

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dOn RTX 4090 · 24 GB
251 LaBSE sentence-transformers · sentence-similarity 470 M 512 ✓ apache-2.0 764 K 2.6 GB
252 w2v-xls-r-uk Yehor · ASR 320 M ✓ apache-2.0 764 K 1.9 GB
253 Ternary-Bonsai-27B-gguf prism-ml · text-generation ✓ apache-2.0 761 K 1.2 GB
254 wav2vec2-xls-r-300m-hebrew imvladikon · ASR 320 M unknown 757 K 1.9 GB
255 Llama-2-7b-chat-hf meta-llama · text-generation · gated 6.7 B ⚠ llama2 753 K 16.3 GB
256 wav2vec2-large-robust-12-ft-emotion-msp-dim audeering · audio-classification 170 M ✗ cc-by-nc-sa-4.0 751 K 1.3 GB
257 Qwen3-0.6B-Base Qwen · text-generation 600 M 33 K ✓ apache-2.0 749 K 1.9 GB
258 yolos-small hustvl · object-detection 30 M ✓ apache-2.0 746 K 0.6 GB
259 Qwen3-30B-A3B-Instruct-2507-AWQ-4bit cyankiwi · text-generation 5.3 B 262 K ✓ apache-2.0 744 K 21.2 GB
260 e5-small-v2 intfloat · sentence-similarity 30 M 512 ✓ mit 739 K 0.7 GB
261 NVIDIA-Nemotron-3-Nano-4B-BF16 nvidia · text-generation 4.0 B 262 K other 739 K 9.8 GB
262 bert-base-multilingual-uncased-sentiment nlptown · text-classification 170 M 512 ✓ mit 735 K 1.3 GB
263 Qwen2.5-Coder-1.5B-Instruct Qwen · text-generation 1.5 B 33 K ✓ apache-2.0 728 K 4.1 GB
264 Qwen3-4B-Base Qwen · text-generation 4.0 B 33 K ✓ apache-2.0 722 K 10.0 GB
265 sd-turbo stabilityai · text-to-image 870 M unknown 722 K 14.9 GB
266 BiRefNet ZhengPeng7 · image-segmentation 220 M ✓ mit 715 K 1.0 GB
267 RMBG-2.0 briaai · image-segmentation · gated 220 M other 702 K 1.5 GB
268 bert-large-cased-finetuned-conll03-english dbmdz · token-classification 330 M 512 unknown 699 K 2.0 GB
269 VibeVoice-ASR microsoft · ASR 8.7 B ✓ mit 695 K 20.9 GB
270 ruri-v3-310m cl-nagoya · sentence-similarity 310 M 8 K ✓ apache-2.0 695 K 1.9 GB
271 nb-wav2vec2-1b-nynorsk NbAiLab · ASR 960 M ✓ apache-2.0 693 K 4.9 GB
272 turn-detector livekit · text-classification 130 M 8 K other 690 K 1.1 GB
273 vit_tiny_patch16_224.augreg_in21k_ft_in1k timm · image-classification 10 M ✓ apache-2.0 689 K 0.5 GB
274 xlm-roberta-base-language-detection papluca · text-classification 280 M 512 ✓ mit 687 K 1.8 GB
275 Mistral-7B-v0.1 mistralai · text-generation 7.2 B 33 K ✓ apache-2.0 679 K 17.5 GB
276 vit_base_patch8_224.augreg2_in21k_ft_in1k timm · image-classification 90 M ✓ apache-2.0 679 K 0.9 GB
277 repvgg_a0.rvgg_in1k timm · image-classification 10 M ✓ mit 669 K 0.5 GB
278 vit-base-nsfw-detector AdamCodd · image-classification 90 M ✓ apache-2.0 660 K 0.9 GB
279 VibeVoice-Realtime-0.5B microsoft · text-to-speech 1.0 B ✓ mit 657 K 2.9 GB
280 bloom-560m bigscience · text-generation 560 M ⚠ bigscience-bloom-rail-1.0 628 K 1.8 GB
281 wav2vec2-large-xlsr-53-gender-recognition-librispeech alefiury · audio-classification 320 M ✓ apache-2.0 626 K 1.9 GB
282 mask2former-swin-large-ade-semantic facebook · image-segmentation 220 M other 625 K 1.5 GB
283 DeepSeek-R1-Distill-Qwen-1.5B deepseek-ai · text-generation 1.8 B 131 K ✓ mit 619 K 4.7 GB
284 OLMo-2-0425-1B allenai · text-generation 1.5 B 4 K ✓ apache-2.0 618 K 7.3 GB
285 parakeet-tdt-0.6b-v3-gguf handy-computer · ASR ✓ cc-by-4.0 606 K 1.0 GB
286 Qwen2-7B-Instruct Qwen · text-generation 7.6 B 33 K ✓ apache-2.0 606 K 18.4 GB
287 Qwen2.5-Coder-7B-Instruct-AWQ Orion-zhen · text-generation 7.6 B 33 K ✓ apache-2.0 604 K 7.9 GB
288 cryptobert ElKulako · text-classification 120 M 512 ✓ mit 602 K 1.1 GB
289 VLM2Vec-Full TIGER-Lab · text-generation 4.2 B 131 K ✓ apache-2.0 599 K 10.2 GB
290 deepseek-coder-7b-instruct-v1.5 deepseek-ai · text-generation 6.9 B 4 K other 599 K 16.7 GB
291 nsfw_image_detector Freepik · image-classification 90 M ✓ mit 591 K 0.7 GB
292 LFM2.5-1.2B-Instruct LiquidAI · text-generation 1.2 B 128 K other 584 K 3.3 GB
293 snowflake-arctic-embed-s Snowflake · sentence-similarity 30 M 512 ✓ apache-2.0 584 K 0.7 GB
294 Qwen3-Coder-30B-A3B-Instruct-AWQ-4bit cyankiwi · text-generation 5.3 B 262 K ✓ apache-2.0 583 K 21.2 GB
295 nb-wav2vec2-1b-bokmaal-v2 NbAiLab · ASR 960 M ✓ apache-2.0 581 K 4.9 GB
296 Qwen3-Reranker-4B-W4A16-G128 boboliu · text-classification 4.1 B 41 K ✓ apache-2.0 578 K 4.0 GB
297 encodec_24khz facebook · feature-extraction 20 M unknown 577 K 0.6 GB
298 gemma-2-2b-it google · text-generation · gated 2.6 B ⚠ gemma 576 K 6.6 GB
299 sentence-bert-base-ja-mean-tokens-v2 sonoisa · feature-extraction 110 M 512 cc-by-sa-4.0 575 K 1.0 GB
300 Qwen3-TTS-12Hz-1.7B-VoiceDesign Qwen · text-to-speech 1.9 B ✓ apache-2.0 574 K 5.8 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.