ReadyArt / text-generation updated 1 month ago

gemma-4-31B-it-scotoma-2-GGUF

#scotoma2{ --bg:#07090d; --panel:#0b1016; --ink:#edf1f4; --soft:#b0bac5; --dim:#7a8591; --faint:#59636f; --line:rgba(150,178,205,.14); --water:#6fc3d8; --water2:#4e95b5; --ember:#e2a24d; background:var(--bg); color:var(--ink); max-width:940px; margin:0 auto; font-family:system-ui,-apple-system,"Segoe UI",Roboto,Helvetica,Arial,sans-serif; font-size:16px; line-height:1.62; border:1px solid var(--line); overflow:hidden...

Params
Context
Downloads 30d
274 K
Likes
38
Commercial use: allowed apache-2.0 Not gated GGUF View on Hugging Face ↗

Download history

daily snapshots · 21 days
274 K183 K
Sep 1Sep 8Sep 15Sep 21

Can you run it?

Estimated VRAM at 8K context unless noted. Pick your hardware to see the verdict per quantization.

FileQuantSizeEst. VRAMVerdict on RTX 4090 · 24 GB
gemma-4-31B-scotoma-2-IQ3_XXS.gguf IQ3_XXS 12.1 GB 13.8 GB ✅ Runs comfortably
gemma-4-31B-scotoma-2-IQ3_XS.gguf IQ3_XS 13.1 GB 14.9 GB ✅ Runs comfortably
gemma-4-31B-scotoma-2-Q3_K_S.gguf Q3_K_S 13.8 GB 15.6 GB ✅ Runs comfortably
gemma-4-31B-scotoma-2-IQ3_M.gguf IQ3_M 14.4 GB 16.4 GB ✅ Runs comfortably
gemma-4-31B-scotoma-2-Q3_K_M.gguf Q3_K_M 15.3 GB 17.3 GB ✅ Runs comfortably
gemma-4-31B-scotoma-2-IQ4_XS.gguf IQ4_XS 16.7 GB 18.9 GB ✅ Runs comfortably
gemma-4-31B-scotoma-2-Q4_K_S.gguf Q4_K_S 17.8 GB 20.0 GB ⚠️ Tight — reduce context
gemma-4-31B-scotoma-2-Q4_K_M.gguf Q4_K_M 18.7 GB 21.1 GB ⚠️ Tight — reduce context
Estimate: file size × 1.1 + KV cache at 8K + 0.5 GB overhead. Not a benchmark — how we calculate this.

Run it

copy-paste, exact tags checked against the Hub
~ · ollama · Q4_K_M
$ ollama run gemma-4-31b-it-scotoma-2-gguf

# pin the quantization explicitly
$ ollama run gemma-4-31b-it-scotoma-2-gguf-q4_k_m
est. VRAM 21.1 GBon RTX 4090 · 24 GBJSON API →

Specifications

Licence
apache-2.0
First seen on the Hub
2026-08-06
Training datasets
undisclosed
Added to our catalog
2026-09-01
Compare with any text-generation model