handy-computer / automatic-speech-recognition updated 1 month ago

nemotron-3.5-asr-streaming-0.6b-gguf

GGUF conversions of nvidia/nemotron-3.5-asr-streaming-0.6b for use with transcribe.cpp.

Params
Context
Downloads 30d
2.0 M
Likes
3
Licence: other Not gated GGUF 28 languages View on Hugging Face ↗

Download history

daily snapshots · 10 days
2.0 M1.8 M
Jul 28Jul 31Aug 3Aug 6

Can you run it?

Estimated VRAM at 8K context unless noted. Pick your hardware to see the verdict per quantization.

FileQuantSizeEst. VRAMVerdict on RTX 4090 · 24 GB
nemotron-3.5-asr-streaming-0.6b-Q4_K_M.gguf Q4_K_M 0.5 GB 1.0 GB ✅ Runs comfortably
nemotron-3.5-asr-streaming-0.6b-Q5_K_M.gguf Q5_K_M 0.6 GB 1.1 GB ✅ Runs comfortably
nemotron-3.5-asr-streaming-0.6b-Q6_K.gguf Q6_K 0.6 GB 1.2 GB ✅ Runs comfortably
nemotron-3.5-asr-streaming-0.6b-Q8_0.gguf Q8_0 0.8 GB 1.3 GB ✅ Runs comfortably
nemotron-3.5-asr-streaming-0.6b-F32.gguf GGUF 2.6 GB 3.3 GB ✅ Runs comfortably
Estimate: file size × 1.1 + KV cache at 8K + 0.5 GB overhead. Not a benchmark — how we calculate this.

Specifications

Licence
other
First seen on the Hub
2026-06-07
Base model
nemotron-3.5-asr-streaming-0.6b
Training datasets
undisclosed
Added to our catalog
2026-07-28

Family

Base model and the most-downloaded derivatives in the catalog.