deepgrove / text-generation updated 1 month ago

maple-preview-GGUF

Variant GGUF size --- ---: TQ10 + Q4K head 4.64 GiB TQ10 + FP16 head 5.06 GiB TQ20 + Q4K head 5.50 GiB TQ20 + FP16 head 5.91 GiB

Params
Context
Downloads 30d
424 K
Likes
63
Commercial use: allowed mit Not gated GGUF 1 languages View on Hugging Face ↗

Download history

daily snapshots · 38 days
▲ 14 K in the last 30 days (3.1%)
446 K315 K
Aug 15Aug 27Sep 9Sep 21

Can you run it?

Estimated VRAM at 8K context unless noted. Pick your hardware to see the verdict per quantization.

FileQuantSizeEst. VRAMVerdict on RTX 4090 · 24 GB
maple-preview-TQ1_0-head-F16.gguf Q1_0 5.4 GB 6.5 GB ✅ Runs comfortably
maple-preview-TQ2_0-head-F16.gguf Q2_0 6.3 GB 7.5 GB ✅ Runs comfortably
Estimate: file size × 1.1 + KV cache at 8K + 0.5 GB overhead. Not a benchmark — how we calculate this.

Run it

copy-paste, exact tags checked against the Hub
~ · ollama · Q1_0
$ ollama run maple-preview-gguf

# pin the quantization explicitly
$ ollama run maple-preview-gguf-q1_0
est. VRAM 6.5 GBon RTX 4090 · 24 GBJSON API →

Specifications

Licence
mit
First seen on the Hub
2026-08-06
Training datasets
undisclosed
Added to our catalog
2026-08-15
Compare with any text-generation model