Downloads · 30 days
0
ZDisket/echolancer-v0.1-base
echolancer-v0.1-base is a text-to-speech model from ZDisket. Use it when you need text read aloud. The card lists the license as mit.
This is a TTS model pretrained on the pre-tokenized Emilia dataset. Since there's no speaker conditioning, the speaker is random at inference. This model has 177M parameters and it was trained from scratch on a single…
Downloads · 30 days
0
Access
Public
Updated Nov 5, 2025
Repo size
2.1 GB
Likes
0
Public
Click a slice to open those files.
.pt2.1 GB · 100%
From the Hugging Face model README
This is a TTS model pretrained on the pre-tokenized Emilia dataset. Since there's no speaker conditioning, the speaker is random at inference. This model has 177M parameters and it was trained from scratch on a single AMD Instinct MI300X for ~2.5 days with the ROCm PyTorch Training v25.7 container.
The training objective was standard next-token prediction on concatenated text-audio tokens.
For more information including a Colab notebook, see the repository.