Qwen3.8-27B-iMatrix-NVFP4-MTP-GGUF vs GLM-5.3-Flash

Specs, VRAM requirements and download trends — updated 20 September 2026.

cdiamond · A
Params
Context
30d
4.0 M
apache-2.0 · commercial OK
zai-org · B
Params
321.3 B
Context
30d
2.9 M
mit · commercial OK

Specification comparison

Differences are highlighted; identical values are muted.

Specification Qwen3.8-27B-iMatrix-NVFP4-MTP-GGUF GLM-5.3-Flash
Parameters 321.3 B
Architecture Glm5NextForConditionalGeneration
Context length
Licence apache-2.0 mit
Commercial use Allowed Allowed
Languages 2
Downloads 30d 4,019,016 2,905,932
Downloads all time 4.0 M 2.9 M
Quantizations on the Hub GGUF SAFETENSORS
Gated No No
First seen on the Hub 2026-08-17 2026-08-25

Download trend

Daily snapshots, last 19 days (2 Sep – 20 Sep)

4.0 M190 K
Qwen3.8-27B-iMatrix-NVFP4-MTP-GGUF GLM-5.3-Flash

VRAM side by side

On RTX 4090 · 24 GB · 8K context unless noted

GLM-5.3-Flash · fp16 409.9 GB / 24 GB
❌ Won’t fit

Adjacent comparisons