wav2vec2-indonesian-javanese-sundanese
This is the model built for the project Multilingual Speech Recognition for Indonesian Languages. It is a fine-tuned facebook/wav2vec2-large-xlsr-53 model on the Indonesian Common Voice dataset, High-quality TTS data for Javanese - SLR41, and High-quality TTS data for Sundanese - SLR44 datasets.
Params
—
Context
—
Downloads 30d
3.1 M
Likes
15
Download history
daily snapshots · 55 days
▲ 1.1 M in the last 30 days (59.0%)
3.7 M1.9 M
Aug 22Sep 1Sep 11Sep 20
3.7 M1.4 M
Jul 28Aug 15Sep 2Sep 20
Specifications
- Architecture
- Wav2Vec2ForCTC
- Vocabulary
- 30
- Layers / heads
- 24 / 16
- Licence
- apache-2.0
- First seen on the Hub
- 2022-03-02
- Training datasets
- mozilla-foundation/common_voice_7_0, openslr, magic_data, titml
- Common Voice 7 (reported)
- 1.577
- Common Voice 6.1 (reported)
- 1.472
- Robust Speech Event - Dev Data (reported)
- 48.94
- Robust Speech Event - Test Data (reported)
- 68.95
- Added to our catalog
- 2026-07-28
Compare with any automatic-speech-recognition model