Downloads · 30 days
126
100% of all-time downloads
computeram/cosyvoice3-badini-tts
cosyvoice3-badini-tts is a text-to-speech model from computeram. Use it when you need text read aloud. It is set up for cosyvoice. The card lists the license as apache-2.0.
A CosyVoice3 model (https://github.com/FunAudioLLM/CosyVoice) fine-tuned for zero-shot text-to-speech in the Badini dialect of Kurdish.
Downloads · 30 days
126
100% of all-time downloads
All-time downloads
126
Public
Repo size
6.7 GB
Likes
0
Public
Click a slice to open those files.
.pt3.4 GB · 51%
From the Hugging Face model README
A CosyVoice3 model (https://github.com/FunAudioLLM/CosyVoice) fine-tuned for zero-shot text-to-speech in the Badini dialect of Kurdish.
import sys
sys.path.append("CosyVoice")
sys.path.append("CosyVoice/third_party/Matcha-TTS")
from cosyvoice.cli.cosyvoice import AutoModel
import torchaudio
cosy = AutoModel(model_dir="<path to downloaded repo>")
prompt_wav = "path/to/prompt.wav"
prompt_text = "transcript of the prompt audio"
instruct = "You are a helpful assistant.<|endofprompt|>"
for out in cosy.inference_zero_shot(
"text to synthesize", instruct + prompt_text, prompt_wav, stream=False
):
torchaudio.save("output.wav", out["tts_speech"], cosy.sample_rate)