Downloads · 30 days
0
Salupala/SonicVox-Multilingual
SonicVox-Multilingual is a text-to-speech model from Salupala. Use it when you need text read aloud. It is set up for chatterbox. The card lists the license as mit.
🎙️ SonicVox-Multilingual is a multilingual Text-to-Speech (TTS) model built on top of Chatterbox Multilingual by Resemble AI.
Downloads · 30 days
0
Access
Public
Updated Aug 17, 2026
Repo size
4.3 GB
Likes
0
Public
Click a slice to open those files.
.safetensors3.2 GB · 75%
From the Hugging Face model README
🎙️ SonicVox-Multilingual is a multilingual Text-to-Speech (TTS) model built on top of Chatterbox Multilingual by Resemble AI.
It is designed to generate natural, expressive speech across multiple languages with support for zero-shot voice cloning using a reference audio sample.
SonicVox-Multilingual is based on the multilingual Chatterbox model.
Supported languages depend on the underlying checkpoint and include multiple languages such as:
Always verify the supported language IDs with the specific Chatterbox checkpoint being used.
SonicVox-Multilingual supports voice cloning using a reference audio file.
Example:
import torchaudio as ta
from chatterbox.mtl_tts import ChatterboxMultilingualTTS
model = ChatterboxMultilingualTTS.from_pretrained(device="cuda")
text = "Welcome to SonicVox Multilingual."
wav = model.generate(
text,
language_id="en",
audio_prompt_path="reference.wav"
)
ta.save("output.wav", wav, model.sr)