CodeMasterCody3D / text-generation updated 5 days ago

taardis-27b-full-ternary

A 27-billion-parameter transformer at 1.75 bits per weight — 5.90 GB — where every weight is a ternary integer {-1, 0, +1} × scale: body, attention, MLP, LM head and embedding table included, with norms and group scales on the integer grid too (balanced-ternary digit stacks). And V2 ships the pipeline's correction system: The Doctors — 496 cross-layer low-rank ternary branches that ride alongside the frozen weights a...

Params
Context
Downloads 30d
268 K
Likes
10
Commercial use: allowed apache-2.0 Not gated GGUF · SAFETENSORS View on Hugging Face ↗

Download history

daily snapshots · 14 days
268 K249 K
Sep 8Sep 12Sep 17Sep 21

Can you run it?

Estimated VRAM at 8K context unless noted. Pick your hardware to see the verdict per quantization.

FileQuantSizeEst. VRAMVerdict on RTX 4090 · 24 GB
TAARDIS-27B-Full-Ternary-V1.gguf GGUF 7.2 GB 8.4 GB ✅ Runs comfortably
model.safetensors fp16 24.4 GB 27.4 GB ❌ Won’t fit
Estimate: file size × 1.1 + KV cache at 8K + 0.5 GB overhead. Not a benchmark — how we calculate this.

Run it

copy-paste, exact tags checked against the Hub
~ · ollama · GGUF
$ ollama run taardis-27b-full-ternary

# pin the quantization explicitly
$ ollama run taardis-27b-full-ternary-gguf
est. VRAM 8.4 GBon RTX 4090 · 24 GBJSON API →

Specifications

Licence
apache-2.0
First seen on the Hub
2026-09-03
Base model
Qwen3.8-27B
Training datasets
undisclosed
Added to our catalog
2026-09-08

Family

Base model and the most-downloaded derivatives in the catalog.

Compare with any text-generation model