Downloads · 30 days
23
2% of all-time downloads
Fastino06/Shona_TTS
Shona_TTS is a text-to-speech model from Fastino06. Use it when you need text read aloud. It is set up for transformers. The card lists the license as cc-by-nc-4.0.
This repository contains the Shona (sna) language text-to-speech (TTS) model checkpoint.
Downloads · 30 days
23
2% of all-time downloads
All-time downloads
1.3K
Public
Parameters
36.3M
291 MB on disk
Likes
1
Public
Click a slice to open those files.
.bin145 MB · 50%
From the Hugging Face model README
This repository contains the Shona (sna) language text-to-speech (TTS) model checkpoint.
pip install --upgrade transformers accelerate
Then, run inference with the following code-snippet:
# Load model directly
from transformers import AutoTokenizer, AutoModelForTextToWaveform
tokenizer = AutoTokenizer.from_pretrained("Fastino06/ff")
model = AutoModelForTextToWaveform.from_pretrained("Fastino06/ff")
text = "some example text in the Shona language"
inputs = tokenizer(text, return_tensors="pt")
with torch.no_grad():
output = model(**inputs).waveform
The resulting waveform can be saved as a .wav file:
import scipy
scipy.io.wavfile.write("fassy.wav", rate=model.config.sampling_rate, data=output)
Or displayed in a Jupyter Notebook / Google Colab:
from IPython.display import Audio
Audio(output, rate=model.config.sampling_rate)
This model was developed by Fastino Mateteva
.