Downloads · 30 days
80
90% of all-time downloads
petrusilius/modus-tts
modus-tts is a text-to-speech model from petrusilius. Use it when you need text read aloud. It is set up for piper. The card lists the license as mit.
Piper TTS voice trained on MODUS dialogue from Fallout 76.
Downloads · 30 days
80
90% of all-time downloads
All-time downloads
89
Public
Repo size
458 MB
Likes
1
Public
Click a slice to open those files.
.onnx457 MB · 100%
From the Hugging Face model README
Piper TTS voice trained on MODUS dialogue from Fallout 76.
| File | Epochs | Training samples |
|---|---|---|
modus_10000.onnx | 10,000 | 421 voice lines |
modus_10000_v2.onnx | 10,000 | 682 voice lines |
Both share the same base, settings and hardware. The voice character is near-identical; v2 is noticeably more robust and slurs less, especially on words outside the training vocabulary. Use v2.
| Property | Value |
|---|---|
| Base checkpoint | en_US-lessac-high |
| Quality | high |
| Sample rate | 22,050 Hz |
| Format | ONNX |
| Language | English |
| Batch size | 8 |
| Precision | 16-bit AMP |
| GPU | NVIDIA RTX A2000 12GB |
| Training time | ~4–5 days per run |
Training vocabulary: 1,984 unique words across 11,117 tokens. Median line length 17 words.
pip install piper-tts
wget https://huggingface.co/petrusilius/modus-tts/resolve/main/modus_10000_v2.onnx
wget https://huggingface.co/petrusilius/modus-tts/resolve/main/modus_10000_v2.onnx.json
Both files must sit in the same folder.
echo "We have you now, General." | \
piper --model modus_10000_v2.onnx --output_file output.wav
Useful flags:
| Flag | Effect |
|---|---|
--length_scale 1.3 | slower |
--length_scale 0.8 | faster |
--sentence_silence 0.5 | longer pause between sentences (default 0.2) |
Plain text only, no SSML.
forty two, not 42General, not Gen.... for mid-sentence pausesNon-commercial fan project. Fallout 76 and all related assets are property of Bethesda Softworks.