1,054 models · refreshed nightly

All models

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dMin VRAM
01 all-MiniLM-L6-v2 sentence-transformers · sentence-similarity 20 M 512 ✓ apache-2.0 254.1 M from 0.6 GB
02 bge-small-en-v1.5 BAAI · feature-extraction 30 M 512 ✓ mit 64.5 M from 0.7 GB
03 paraphrase-multilingual-MiniLM-L12-v2 sentence-transformers · sentence-similarity 120 M 512 ✓ apache-2.0 45.9 M from 1.0 GB
04 bge-m3 BAAI · sentence-similarity 8 K ✓ mit 38.0 M
05 t5-small google-t5 · translation 60 M ✓ apache-2.0 24.9 M from 0.8 GB
06 Qwen3-0.6B Qwen · text-generation 750 M 41 K ✓ apache-2.0 23.0 M from 2.3 GB
07 all-mpnet-base-v2 sentence-transformers · sentence-similarity 110 M 512 ✓ apache-2.0 22.6 M from 1.0 GB
08 Qwen3-VL-8B-Instruct Qwen · image-text-to-text 8.8 B ✓ apache-2.0 19.5 M from 21.1 GB
09 mobilenetv3_small_100.lamb_in1k timm · image-classification <0.1 M ✓ apache-2.0 18.0 M from 0.5 GB
10 wav2vec2-large-xlsr-53-japanese jonatasgrosman · ASR ✓ apache-2.0 17.9 M
11 bge-reranker-v2-m3 BAAI · text-classification 570 M 8 K ✓ apache-2.0 17.6 M from 3.1 GB
12 gpt2 openai-community · text-generation 140 M ✓ mit 15.4 M from 1.1 GB
13 nomic-embed-text-v1.5 nomic-ai · sentence-similarity 140 M 2 K ✓ apache-2.0 14.7 M from 1.1 GB
14 Qwen3-8B Qwen · text-generation 8.2 B 41 K ✓ apache-2.0 12.9 M from 19.7 GB
15 Qwen3-Coder-30B-A3B-Instruct-GGUF unsloth · text-generation ✓ apache-2.0 12.7 M from 12.9 GB
16 efficientnet_b3.ra2_in1k timm · image-classification 10 M ✓ apache-2.0 12.7 M from 0.6 GB
17 multilingual-e5-small intfloat · sentence-similarity 120 M 512 ✓ mit 12.3 M from 1.0 GB
18 Kokoro-82M hexgrad · text-to-speech 82 M ✓ apache-2.0 11.6 M from 0.7 GB
19 bge-large-en-v1.5 BAAI · feature-extraction 340 M 512 ✓ mit 11.4 M from 2.0 GB
20 whisperkit-coreml argmaxinc · ASR ✓ mit 11.2 M
21 bge-base-en-v1.5 BAAI · feature-extraction 110 M 512 ✓ mit 10.5 M from 1.0 GB
22 gemma-4-26B-A4B-it google · image-text-to-text 26.5 B ✓ apache-2.0 9.9 M from 61.3 GB
23 paraphrase-multilingual-mpnet-base-v2 sentence-transformers · sentence-similarity 280 M 512 ✓ apache-2.0 9.9 M from 1.8 GB
24 Qwen3.6-35B-A3B-FP8 Qwen · image-text-to-text 36.0 B ✓ apache-2.0 9.8 M from 47.1 GB
25 Qwen2.5-7B-Instruct Qwen · text-generation 7.6 B 33 K ✓ apache-2.0 9.7 M from 18.4 GB
26 Qwen3.5-9B Qwen · image-text-to-text 9.7 B ✓ apache-2.0 9.3 M from 23.2 GB
27 gemma-4-31B-it google · image-text-to-text 32.7 B ✓ apache-2.0 9.1 M from 74.2 GB
28 Qwen3-Embedding-0.6B Qwen · feature-extraction 600 M 33 K ✓ apache-2.0 8.6 M from 1.9 GB
29 Qwen2.5-0.5B-Instruct Qwen · text-generation 490 M 33 K ✓ apache-2.0 8.5 M from 1.7 GB
30 Qwen3.6-35B-A3B-NVFP4 nvidia · text-generation 18.7 B ✓ apache-2.0 8.3 M from 29.1 GB
31 clap-htsat-fused laion · audio-classification 150 M ✓ apache-2.0 8.2 M from 1.2 GB
32 speaker-diarization-3.1 pyannote · ASR · gated ✓ mit 8.2 M
33 multilingual-e5-base intfloat · sentence-similarity 280 M 512 ✓ mit 7.6 M from 1.8 GB
34 opt-125m facebook · text-generation 125 M 2 K other 7.4 M from 0.8 GB
35 Qwen3.8-27B Qwen · image-text-to-text 27.8 B ✓ apache-2.0 7.4 M from 65.8 GB
36 XTTS-v2 coqui · text-to-speech other 7.3 M
37 Qwen2.5-1.5B-Instruct Qwen · text-generation 1.5 B 33 K ✓ apache-2.0 7.2 M from 4.1 GB
38 vit-base-patch16-224 google · image-classification 90 M ✓ apache-2.0 7.2 M from 0.9 GB
39 multilingual-e5-large intfloat · feature-extraction 560 M 512 ✓ mit 7.2 M from 3.0 GB
40 Qwen3-4B Qwen · text-generation 4.0 B 41 K ✓ apache-2.0 7.1 M from 10.0 GB
41 Qwen3.8-27B-FP8 Qwen · image-text-to-text 27.8 B ✓ apache-2.0 7.1 M from 38.6 GB
42 Llama-3.2-1B-Instruct meta-llama · text-generation · gated 1.2 B ⚠ llama3.2 7.0 M from 3.4 GB
43 Qwen3.5-4B Qwen · image-text-to-text 4.7 B ✓ apache-2.0 6.9 M from 11.5 GB
44 OTel-2.0-LLM-31B-IT farbodtavakkoli · text-generation 31.3 B 262 K ✓ apache-2.0 6.9 M from 144.6 GB
45 Qwen2.5-VL-7B-Instruct Qwen · image-text-to-text 8.3 B 128 K ✓ apache-2.0 6.9 M from 20.0 GB
46 whisper-large-v3-turbo openai · ASR 810 M ✓ mit 6.8 M from 2.4 GB
47 gpt-oss-20b openai · text-generation 21.5 B 131 K ✓ apache-2.0 6.7 M from 34.0 GB
48 granite-embedding-small-english-r2 ibm-granite · feature-extraction 50 M 8 K ✓ apache-2.0 6.3 M from 0.6 GB
49 wav2vec2-large-xlsr-53-portuguese jonatasgrosman · ASR ✓ apache-2.0 6.0 M
50 Llama-3.1-8B-Instruct meta-llama · text-generation · gated 8.0 B ⚠ llama3.1 5.9 M from 19.4 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.