GLM-4.7-Flash vs Qwen3.6-27B-NVFP4
Specs, VRAM requirements and download trends — updated 2 August 2026.
Specification comparison
Differences are highlighted; identical values are muted.
| Specification | GLM-4.7-Flash | Qwen3.6-27B-NVFP4 |
|---|---|---|
| Parameters | 31.2 B | 18.2 B |
| Architecture | Glm4MoeLiteForCausalLM | Qwen3_5ForConditionalGeneration |
| Context length | 202,752 | — |
| Licence | mit | apache-2.0 |
| Commercial use | Allowed | Allowed |
| Languages | 2 | — |
| Downloads 30d | 1,852,678 | 217,456 |
| Downloads all time | 13.7 M | 3.2 M |
| Quantizations on the Hub | SAFETENSORS | SAFETENSORS |
| Gated | No | No |
| First seen on the Hub | 2026-01-19 | 2026-06-22 |
Download trend
Daily snapshots, last 56 days (28 Jul – 21 Sep)
2.2 M217 K
GLM-4.7-Flash
Qwen3.6-27B-NVFP4
A overtook B on 1 Aug
VRAM side by side
On RTX 4090 · 24 GB · 8K context unless noted
Qwen3.6-27B-NVFP4 · fp16
27.3 GB / 24 GB
GLM-4.7-Flash · fp16 @ 198K ctx
185.1 GB / 24 GB