privacy-filter-nemotron-GGUF
GGUF conversion of OpenMed/privacy-filter-nemotron, a fine-grained PII token-classification model — a fine-tune of openai/privacy-filter on the nvidia/Nemotron-PII dataset. It labels every token with a BIOES tag over 55 PII categories (221 classes) in a single forward pass, then decodes coherent spans with a constrained Viterbi procedure — so it can be served locally with no Python as the encoder/NER tier of a PII re...
Params
—
Context
—
Downloads 30d
951 K
Likes
0
Download history
daily snapshots · 38 days
▲ 458 K in the last 30 days (92.8%)
963 K527 K
Aug 22Sep 1Sep 11Sep 20
963 K234 K
Aug 14Aug 26Sep 8Sep 20
Can you run it?
Estimated VRAM at 8K context unless noted. Pick your hardware to see the verdict per quantization.
| File | Quant | Size | Est. VRAM | Verdict on RTX 4090 · 24 GB |
|---|---|---|---|---|
| privacy-filter-nemotron-q8.gguf | Q8 | 1.6 GB | 2.3 GB | ✅ Runs comfortably |
| privacy-filter-nemotron-f16.gguf | GGUF | 2.8 GB | 3.6 GB | ✅ Runs comfortably |
Estimate: file size × 1.1 + KV cache at 8K + 0.5 GB overhead. Not a benchmark — how we calculate this.
Specifications
- Licence
- apache-2.0
- First seen on the Hub
- 2026-06-19
- Training datasets
- nvidia/Nemotron-PII
- Added to our catalog
- 2026-08-14
Compare with any token-classification model