Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF
Important: This is the first fine tune to exceed 700 "arc-c" (The OpenAI, Claude and Gemini "zone of intelligence") in both 8 bit and 4 bit. This repo contains both "regular" and "MTP" Neo MAX Imatrix GGUF quants. Many other additional quant types avail too. 3rd parties confirm this model's performance in the "community tab". (40B version has entered Beta)
Params
—
Context
—
Downloads 30d
1.6 M
Likes
1,593
Download history
daily snapshots · 7 days1.6 M956 K
Jul 31Aug 2Aug 4Aug 6
Can you run it?
Estimated VRAM at 8K context unless noted. Pick your hardware to see the verdict per quantization.
| File | Quant | Size | Est. VRAM | Verdict on RTX 4090 · 24 GB |
|---|---|---|---|---|
| Qwen3.6-27B-Fable-Fus-711-UnHeretic-NM-DAU-NEO-MAX-NEO-MTP-IQ2_M.gguf | IQ2_M | 12.1 GB | 13.8 GB | ✅ Runs comfortably |
| Qwen3.6-27B-Fable-Fus-711-UnHeretic-NM-DAU-NEO-MAX-NEO-MTP-IQ3_M.gguf | IQ3_M | 14.5 GB | 16.5 GB | ✅ Runs comfortably |
| Qwen3.6-27B-Fable-Fus-711-UnHeretic-NM-DAU-NEO-MAX-NEO-MTP-IQ4_XS.gguf | IQ4_XS | 17.0 GB | 19.2 GB | ✅ Runs comfortably |
| Qwen3.6-27B-Fable-Fus-711-UnHeretic-NM-DAU-NEO-MAX-NEO-MTP-Q4_K_S.gguf | Q4_K_S | 17.5 GB | 19.8 GB | ⚠️ Tight — reduce context |
| Qwen3.6-27B-Fable-Fus-711-UnHeretic-NM-DAU-NEO-MAX-NEO-MTP-IQ4_NL.gguf | IQ4_NL | 17.8 GB | 20.0 GB | ⚠️ Tight — reduce context |
| Qwen3.6-27B-Fable-Fus-711-UnHeretic-NM-DAU-NEO-MAX-NEO-MTP-Q4_K_M.gguf | Q4_K_M | 18.5 GB | 20.8 GB | ⚠️ Tight — reduce context |
| Qwen3.6-27B-Fable-Fus-711-UnHeretic-NM-DAU-NEO-MAX-NEO-MTP-Q5_K_M.gguf | Q5_K_M | 21.2 GB | 23.8 GB | ⚠️ Tight — reduce context |
| Qwen3.6-27B-Fable-Fus-711-UnHeretic-NM-DAU-NEO-MAX-NEO-AMD-MTP-Q6_K.gguf | Q6_K | 24.0 GB | 26.9 GB | ❌ Won’t fit |
Estimate: file size × 1.1 + KV cache at 8K + 0.5 GB overhead. Not a benchmark — how we calculate this.
Specifications
- Licence
- apache-2.0
- First seen on the Hub
- 2026-07-17
- Training datasets
- DavidAU/Polar-STRICT-Datasets, DavidAU/F451-STRICT-Datasets
- Added to our catalog
- 2026-07-31
Compare with
Sponsored · GPU cloud
Not enough VRAM?
Spin up a 24 GB L4 instance in 40 seconds. $0.44/hr.