unsloth / image-text-to-text updated 5 months ago

Qwen3.6-35B-A3B-GGUF

See Unsloth Dynamic 2.0 GGUFs for our quantization benchmarks. NEW: Developer Role Support so Qwen3.6 can work in Codex, OpenCode and more! Qwen3.6 can now be run and fine-tuned in Unsloth Studio. Read our guide. Tool calling improvements: Makes parsing nested objects to make tool calling succeed more. Example of Qwen3.6 (4-bit GGUF) running in Unsloth Studio with tool-calling:

Params
Context
Downloads 30d
1.3 M
Likes
1,618
Commercial use: allowed apache-2.0 Not gated GGUF View on Hugging Face ↗

Download history

daily snapshots · 51 days
▲ 5 K in the last 30 days (0.4%)
1.3 M848 K
Aug 2Aug 19Sep 5Sep 21

Can you run it?

Estimated VRAM at 8K context unless noted. Pick your hardware to see the verdict per quantization.

FileQuantSizeEst. VRAMVerdict on RTX 4090 · 24 GB
Qwen3.6-35B-A3B-UD-IQ1_M.gguf IQ1_M 10.0 GB 11.6 GB ✅ Runs comfortably
Qwen3.6-35B-A3B-UD-IQ2_XXS.gguf IQ2_XXS 10.8 GB 12.3 GB ✅ Runs comfortably
Qwen3.6-35B-A3B-UD-IQ2_M.gguf IQ2_M 11.5 GB 13.2 GB ✅ Runs comfortably
Qwen3.6-35B-A3B-UD-IQ3_XXS.gguf IQ3_XXS 13.2 GB 15.0 GB ✅ Runs comfortably
Qwen3.6-35B-A3B-UD-IQ3_S.gguf IQ3_S 13.7 GB 15.5 GB ✅ Runs comfortably
Qwen3.6-35B-A3B-UD-IQ4_NL.gguf IQ4_NL 18.0 GB 20.3 GB ⚠️ Tight — reduce context
Qwen3.6-35B-A3B-Q8_0.gguf Q8_0 36.9 GB 41.1 GB ❌ Won’t fit
Qwen3.6-35B-A3B-BF16-00001-of-00002.gguf GGUF 49.9 GB 55.4 GB ❌ Won’t fit
Estimate: file size × 1.1 + KV cache at 8K + 0.5 GB overhead. Not a benchmark — how we calculate this.

Specifications

Licence
apache-2.0
First seen on the Hub
2026-04-16
Base model
Qwen3.6-35B-A3B
Training datasets
undisclosed
Added to our catalog
2026-08-02

Family

Base model and the most-downloaded derivatives in the catalog.

Compare with any image-text-to-text model