coqui / text-to-speech updated 2 years ago

XTTS-v2

ⓍTTS is a Voice generation model that lets you clone voices into different languages by using just a quick 6-second audio clip. There is no need for an excessive amount of training data that spans countless hours.

Params
Context
Downloads 30d
7.3 M
Likes
3,808
Licence: other Not gated View on Hugging Face ↗

Download history

daily snapshots · 55 days
▲ 1.2 M in the last 30 days (14.4%)
9.3 M7.2 M
Jul 28Aug 15Sep 2Sep 20

Specifications

Licence
other
First seen on the Hub
2023-10-31
Training datasets
undisclosed
Added to our catalog
2026-07-28
Compare with any text-to-speech model