audio-cpp / text-to-speech updated 3 days ago

audio.cpp-gguf

This directory contains audio.cpp-native GGUF conversions of multiple speech models. These files are intended for use with audio.cpp.

Params
Context
Downloads 30d
4.3 M
Likes
157
Licence: other Not gated GGUF · SAFETENSORS View on Hugging Face ↗

Download history

daily snapshots · 34 days
▲ 3.6 M in the last 30 days (533.4%)
4.3 M419 K
Aug 18Aug 29Sep 9Sep 20

Can you run it?

Estimated VRAM at 8K context unless noted. Pick your hardware to see the verdict per quantization.

FileQuantSizeEst. VRAMVerdict on RTX 4090 · 24 GB
model.safetensors fp16 0.8 GB 1.4 GB ✅ Runs comfortably
cohere-transcribe-03-2026-q4_0.gguf Q4_0 1.5 GB 2.2 GB ✅ Runs comfortably
qwen3-tts-12hz-1.7b-base-q8_0_v2.gguf Q8_0_V2 2.7 GB 3.5 GB ✅ Runs comfortably
personaplex-7b-v1-q4_k.gguf Q4_K 7.9 GB 9.1 GB ✅ Runs comfortably
dit_int8.gguf Q4 17.4 GB 19.7 GB ⚠️ Tight — reduce context
dramabox-q8_0.gguf Q8_0 18.9 GB 21.3 GB ⚠️ Tight — reduce context
firered-audio-orig.gguf GGUF 25.0 GB 28.0 GB ❌ Won’t fit
Estimate: file size × 1.1 + KV cache at 8K + 0.5 GB overhead. Not a benchmark — how we calculate this.

Specifications

Licence
other
First seen on the Hub
2026-07-14
Training datasets
undisclosed
Added to our catalog
2026-08-18
Compare with any text-to-speech model