facebook / automatic-speech-recognition updated 4 years ago

wav2vec2-large-960h-lv60-self

The large model pretrained and fine-tuned on 960 hours of Libri-Light and Librispeech on 16kHz sampled speech audio. Model was trained with Self-Training objective. When using the model make sure that your speech input is also sampled at 16Khz.

Params
Context
Downloads 30d
412 K
Likes
162
Commercial use: allowed apache-2.0 Not gated 1 languages View on Hugging Face ↗

Download history

daily snapshots · 10 days
630 K412 K
Jul 28Jul 31Aug 3Aug 6

Specifications

Architecture
Wav2Vec2ForCTC
Vocabulary
32
Layers / heads
24 / 16
Licence
apache-2.0
First seen on the Hub
2022-03-02
Training datasets
librispeech_asr
LibriSpeech (clean) (reported)
1.9
LibriSpeech (other) (reported)
3.9
Added to our catalog
2026-07-28