Downloads · 30 days
129
2% of all-time downloads
fishaudio/fish-speech-1.2
fish-speech-1.2 is a text-to-speech model from fishaudio. Use it when you need text read aloud. It is set up for transformers. The card lists the license as cc-by-nc-sa-4.0.
Fish Speech V1.2 is a leading text-to-speech (TTS) model trained on 300k hours of English, Chinese, and Japanese audio data.
Downloads · 30 days
129
2% of all-time downloads
All-time downloads
7.1K
Public
Repo size
1.1 GB
Likes
210
Public
Click a slice to open those files.
.pth1.1 GB · 100%
From the Hugging Face model README
Fish Speech V1.2 is a leading text-to-speech (TTS) model trained on 300k hours of English, Chinese, and Japanese audio data.
Please refer to Fish Speech Github for more info.
Demo available at Fish Audio.
If you found this repository useful, please consider citing this work:
@misc{fish-speech-v1,
author = {Shijia Liao, Tianyu Li},
title = {Fish Speech V1},
year = {2024},
publisher = {GitHub},
journal = {GitHub repository},
howpublished = {\url{https://github.com/fishaudio/fish-speech}}
}
This model is permissively licensed under the BY-CC-NC-SA-4.0 license. The source code is released under BSD-3-Clause license.