wav2vec2-large-xlsr-53-gender-recognition-librispeech vs ast-finetuned-audioset-10-10-0.4593

Specs, VRAM requirements and download trends — updated 6 August 2026.

alefiury · A
Params
320 M
Context
30d
626 K
apache-2.0 · commercial OK
MIT · B
Params
90 M
Context
30d
1.2 M
bsd-3-clause · commercial OK

Specification comparison

Differences are highlighted; identical values are muted.

Specification wav2vec2-large-xlsr-53-gender-recognition-librispeech ast-finetuned-audioset-10-10-0.4593
Parameters 320 M 90 M
Architecture Wav2Vec2ForSequenceClassification ASTForAudioClassification
Context length
Licence apache-2.0 bsd-3-clause
Commercial use Allowed Allowed
Languages
Downloads 30d 626,341 1,181,816
Downloads all time 7.9 M 2.2 B
Quantizations on the Hub SAFETENSORS SAFETENSORS
Gated No No
First seen on the Hub 2023-04-24 2022-11-14

Download trend

Daily snapshots, last 10 days (28 Jul – 6 Aug)

1.2 M620 K
wav2vec2-large-xlsr-53-gender-recognition-librispeech ast-finetuned-audioset-10-10-0.4593

VRAM side by side

On RTX 4090 · 24 GB · 8K context unless noted

ast-finetuned-audioset-10-10-0.4593 · fp16 0.9 GB / 24 GB
✅ Runs comfortably
wav2vec2-large-xlsr-53-gender-recognition-librispeech · fp16 1.9 GB / 24 GB
✅ Runs comfortably

Adjacent comparisons