Downloads · 30 days
13
2% of all-time downloads
ntdgo/ttsvi
ttsvi is a text-to-speech model from ntdgo. Use it when you need text read aloud. It is set up for transformers. The card lists the license as other.
viⓍTTS là mô hình tạo sinh giọng nói cho phép bạn sao chép giọng nói sang các ngôn ngữ khác nhau chỉ bằng cách sử dụng một đoạn âm thanh nhanh dài 6 giây. Mô hình này được tiếp tục đào tạo từ mô hình XTTS-v2.0.3 bằng…
Downloads · 30 days
13
2% of all-time downloads
All-time downloads
574
Public
Repo size
1.9 GB
Likes
12
Public
Click a slice to open those files.
.pth1.9 GB · 100%
From the Hugging Face model README
viⓍTTS là mô hình tạo sinh giọng nói cho phép bạn sao chép giọng nói sang các ngôn ngữ khác nhau chỉ bằng cách sử dụng một đoạn âm thanh nhanh dài 6 giây. Mô hình này được tiếp tục đào tạo từ mô hình XTTS-v2.0.3 bằng cách mở rộng tokenizer sang tiếng Việt và huấn luyện trên tập dữ liệu viVoice.
viⓍTTS is a voice generation model that lets you clone voices into different languages by using just a quick 6-second audio clip. This model is fine-tuned from the XTTS-v2.0.3 model by expanding the tokenizer to Vietnamese and fine-tuning on the viVoice dataset.
viXTTS supports 18 languages: English (en), Spanish (es), French (fr), German (de), Italian (it), Portuguese (pt), Polish (pl), Turkish (tr), Russian (ru), Dutch (nl), Czech (cs), Arabic (ar), Chinese (zh-cn), Japanese (ja), Hungarian (hu), Korean (ko) Hindi (hi), Vietnamese (vi).
Please checkout this repo
For a quick usage, please checkout this notebook
This model is licensed under Coqui Public Model License.
Fine-tuned by Thinh Le at FPT University HCMC, as a component of Non La's graduation thesis. Contact: