fullstop-punctuation-multilang-large
This model predicts the punctuation of English, Italian, French and German texts. We developed it to restore the punctuation of transcribed spoken language.
Params
560 M
Context
512
Downloads 30d
232 K
Likes
180
Download history
daily snapshots · 57 days
▲ 471 in the last 30 days (0.2%)
232 K232 K
Aug 24Sep 3Sep 13Sep 22
643 K232 K
Jul 28Aug 16Sep 4Sep 22
Can you run it?
Estimated VRAM at 8K context unless noted. Pick your hardware to see the verdict per quantization.
| File | Quant | Size | Est. VRAM | Verdict on RTX 4090 · 24 GB |
|---|---|---|---|---|
| model.safetensors | f32 | 2.2 GB | 3.0 GB | ✅ Runs comfortably |
Estimate: file size × 1.1 + KV cache at 8K + 0.5 GB overhead. Not a benchmark — how we calculate this.
Specifications
- Architecture
- XLMRobertaForTokenClassification
- Parameters
- 560 M
- Tensor type
- F32
- Context length
- 512
- Vocabulary
- 250,002
- Layers / heads
- 24 / 16
- Licence
- mit
- First seen on the Hub
- 2022-03-02
- Training datasets
- wmt/europarl
- Added to our catalog
- 2026-07-28
Compare with any token-classification model