gvs / automatic-speech-recognition updated 5 years ago

wav2vec2-large-xlsr-malayalam

Fine-tuned facebook/wav2vec2-large-xlsr-53 on ml (Malayalam) using the Indic TTS Malayalam Speech Corpus (via Kaggle), Openslr Malayalam Speech Corpus, SMC Malayalam Speech Corpus and IIIT-H Indic Speech Databases. The notebooks used to train model are available here. When using this model, make sure that your speech input is sampled at 16kHz.

Params
Context
Downloads 30d
649 K
Likes
7
Commercial use: allowed apache-2.0 Not gated 1 languages View on Hugging Face ↗

Download history

daily snapshots · 51 days
▲ 265 K in the last 30 days (69.3%)
787 K301 K
Aug 2Aug 19Sep 5Sep 21

Specifications

Architecture
Wav2Vec2ForCTC
Vocabulary
76
Layers / heads
24 / 16
Licence
apache-2.0
First seen on the Hub
2022-03-02
Training datasets
undisclosed
Test split of combined dataset using all datasets mentioned above (reported)
28.43
Added to our catalog
2026-08-02
Compare with any automatic-speech-recognition model