GLM-5.3-Flash vs Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF

Specs, VRAM requirements and download trends — updated 20 September 2026.

zai-org · A
Params
321.3 B
Context
30d
2.9 M
mit · commercial OK
DavidAU · B
Params
Context
30d
1.5 M
apache-2.0 · commercial OK

Specification comparison

Differences are highlighted; identical values are muted.

Specification GLM-5.3-Flash Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF
Parameters 321.3 B
Architecture Glm5NextForConditionalGeneration
Context length
Licence mit apache-2.0
Commercial use Allowed Allowed
Languages 2 2
Downloads 30d 2,905,932 1,483,343
Downloads all time 2.9 M 2.3 M
Quantizations on the Hub SAFETENSORS GGUF
Gated No No
First seen on the Hub 2026-08-25 2026-07-19

Download trend

Daily snapshots, last 22 days (30 Aug – 20 Sep)

2.9 M190 K
GLM-5.3-Flash Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF A overtook B on 12 Sep

VRAM side by side

On RTX 4090 · 24 GB · 8K context unless noted

Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF · Q4_K_M 8.2 GB / 24 GB
✅ Runs comfortably
GLM-5.3-Flash · fp16 409.9 GB / 24 GB
❌ Won’t fit

Adjacent comparisons