Downloads · 30 days
0
ZDisket/echolancer-stage2-base
echolancer-stage2-base is a text-to-speech model from ZDisket. Use it when you need text read aloud. The card lists the license as mit.
This is a TTS model pretrained on the pre-tokenized Emilia dataset. Since there's no speaker conditioning, the speaker is random at inference. This model has 550M parameters and it was trained from scratch on a single…
Downloads · 30 days
0
Access
Public
Updated Nov 15, 2025
Repo size
6.6 GB
Likes
0
Public
Click a slice to open those files.
.pt6.6 GB · 100%
From the Hugging Face model README
This is a TTS model pretrained on the pre-tokenized Emilia dataset. Since there's no speaker conditioning, the speaker is random at inference. This model has 550M parameters and it was trained from scratch on a single AMD Instinct MI300X for ~4 days with the ROCm PyTorch Training v25.7 container.
The training objective was standard next-token prediction on concatenated text-audio tokens.
For more information including a Colab notebook, see the repository.