1,058 models · refreshed nightly

Models that run on Mac M3 · 24 GB

Every model in the catalog with its licence, estimated VRAM and daily-tracked downloads. Filters update the URL — share any view.

#ModelParamsContextCommercial use30dOn Mac M3 · 24 GB
351 Llama-2-7b-chat-hf meta-llama · text-generation · gated 6.7 B ⚠ llama2 447 K 16.3 GB
352 mask2former-swin-large-cityscapes-semantic facebook · image-segmentation 220 M other 444 K 1.5 GB
353 zeta-2.1-autoround-W4A16 LeaderboardModel1 · text-generation 2.2 B 33 K unknown 444 K 7.6 GB
354 msmarco-bert-base-dot-v5 sentence-transformers · sentence-similarity 110 M 512 unknown 444 K 1.0 GB
355 Voxtral-Mini-4B-Realtime-2602-gguf handy-computer · ASR ✓ apache-2.0 442 K 3.6 GB
356 Bonsai-27B-gguf prism-ml · text-generation ✓ apache-2.0 440 K 1.2 GB
357 BioLORD-2023 FremyCompany · sentence-similarity 110 M 512 other 437 K 1.0 GB
358 bge-small-en-v1.5 Xenova · feature-extraction 30 M 512 unknown 433 K 0.6 GB
359 DeepSeek-R1-Distill-Qwen-1.5B deepseek-ai · text-generation 1.8 B 131 K ✓ mit 431 K 4.7 GB
360 vit_small_patch16_224.augreg_in21k_ft_in1k timm · image-classification 20 M ✓ apache-2.0 430 K 0.6 GB
361 bert-base-turkish-cased-mean-nli-stsb-tr emrecan · sentence-similarity 110 M 512 ✓ apache-2.0 429 K 1.0 GB
362 whisper-large-v3-turbo-german primeline · ASR 810 M ✓ apache-2.0 428 K 2.4 GB
363 Qwen3-4B-GGUF Qwen · text-generation ✓ apache-2.0 428 K 3.2 GB
364 wav2vec2-large-mms-1b-azerbaijani-common_voice15.0 nijatzeynalov · ASR 960 M ✗ cc-by-nc-4.0 427 K 4.9 GB
365 Qwen2.5-3B-Instruct-AWQ Qwen · text-generation 3.4 B 33 K other 427 K 4.0 GB
366 llama-nemotron-embed-1b-v2 nvidia · feature-extraction 1.2 B 131 K other 425 K 3.4 GB
367 maple-preview-GGUF deepgrove · text-generation ✓ mit 424 K 6.5 GB
368 Qwen-AgentWorld-35B-A3B-GGUF unsloth · text-generation ✓ apache-2.0 423 K 13.1 GB
369 Bangla-twoclass-Sentiment-Analyzer Arunavaonly · text-classification 280 M 512 ✓ mit 422 K 1.8 GB
370 MiniCPM5-2B openbmb · text-generation 2.5 B 131 K ✓ apache-2.0 421 K 6.4 GB
371 Qwen3-4B-Thinking-2507 Qwen · text-generation 4.0 B 262 K ✓ apache-2.0 420 K 10.0 GB
372 wav2vec2-large-xls-r-300m-bg-d2 DrishtiSharma · ASR 300 M ✓ apache-2.0 419 K 1.2 GB
373 nemotron-speech-streaming-en-0.6b nvidia · ASR 618 M other 419 K 1.4 GB
374 Qwen3.6-14B-A3B-FableVibes-GGUF tvall43 · text-generation ✓ apache-2.0 419 K 6.4 GB
375 all-MiniLM-L6-v2-GGUF leliuga · sentence-similarity 20 M 512 ✓ apache-2.0 416 K 0.5 GB
376 Mistral-7B-v0.1 mistralai · text-generation 7.2 B 33 K ✓ apache-2.0 415 K 17.5 GB
377 vram-16 unslothai · feature-extraction unknown 410 K 0.5 GB
378 Phi-3-mini-4k-instruct microsoft · text-generation 3.8 B 4 K ✓ mit 407 K 9.5 GB
379 S-PubMedBert-MedQuAD TimKond · sentence-similarity 110 M 512 ✓ mit 406 K 1.0 GB
380 Qwen3-8B-GGUF Qwen · text-generation ✓ apache-2.0 404 K 6.0 GB
381 gemma-2-9b-it-AWQ-INT4 hugging-quants · text-generation 9.2 B 8 K ⚠ gemma 402 K 8.7 GB
382 wav2vec2-large-xlsr-marathi sumedh · ASR 320 M ✓ apache-2.0 399 K 1.9 GB
383 Qwen2.5-VL-7B-Instruct-NVFP4 nvidia · text-generation 5.0 B 128 K other 399 K 9.2 GB
384 jina-reranker-m0 jinaai · text-classification 2.4 B 33 K ✗ cc-by-nc-4.0 395 K 6.2 GB
385 gemma-3-270m google · text-generation · gated 270 M ⚠ gemma 395 K 1.1 GB
386 granite-4.1-3b ibm-granite · text-generation 3.4 B 131 K ✓ apache-2.0 393 K 8.5 GB
387 Ternary-Bonsai-8B-gguf prism-ml · text-generation ✓ apache-2.0 392 K 2.9 GB
388 Qwen3.8-27B-DFlash2 incoai · text-generation 1.9 B 262 K ✓ apache-2.0 391 K 5.0 GB
389 Kwaipilot_KAT-Coder-V2.5-Dev-GGUF bartowski · text-generation ✓ apache-2.0 390 K 11.3 GB
390 Phi-4-mini-instruct microsoft · text-generation 3.8 B 131 K ✓ mit 384 K 9.5 GB
391 TwIL-LM3 webAI-Official · text-generation 3.1 B 66 K other 384 K 3.1 GB
392 rtdetr_r50vd_coco_o365 PekingU · object-detection 40 M ✓ apache-2.0 380 K 0.7 GB
393 falcon-7b tiiuae · text-generation 7.2 B ✓ apache-2.0 379 K 17.5 GB
394 pythia-14m EleutherAI · text-generation 10 M 2 K ✓ apache-2.0 378 K 0.5 GB
395 mistral-7b-v0.3-bnb-4bit unsloth · text-generation 7.5 B 33 K ✓ apache-2.0 377 K 6.2 GB
396 Ornith-1.5-9B-OBLITERATED OBLITERATUS · text-generation 9.7 B ✓ mit 371 K 6.3 GB
397 segformer-b0-finetuned-ade-512-512 nvidia · image-segmentation <0.1 M other 368 K 0.5 GB
398 Llama-3.2-3B meta-llama · text-generation · gated 3.2 B ⚠ llama3.2 368 K 8.0 GB
399 Parable-Granite-4.1-3B-Claude-Fable-5-GGUF AnkitAI · text-generation ✓ apache-2.0 366 K 2.8 GB
400 privacy-filter-multilingual-GGUF LocalAI-io · token-classification ✓ apache-2.0 366 K 2.3 GB
VRAM figures are estimates for the smallest available quantization at 8K context — see /methodology. Downloads refresh nightly from the Hugging Face API.