Downloads · 30 days
24
67% of all-time downloads
RootAccess4Life/cond-id-indextts-1.5
cond-id-indextts-1.5 is a text-to-speech model from RootAccess4Life. Use it when you need text read aloud. The card lists the license as other.
An unmodified, commit-pinned mirror of IndexTeam/IndexTTS-1.5, bundled with the matching index-tts source tree, so the cond-ID paper's IndexTTS results stay reproducible.
Downloads · 30 days
24
67% of all-time downloads
All-time downloads
36
Public
Repo size
3.6 GB
Likes
0
Public
Click a slice to open those files.
.pth3.6 GB · 100%
From the Hugging Face model README
An unmodified, commit-pinned mirror of
IndexTeam/IndexTTS-1.5, bundled with the
matching index-tts source tree, so the cond-ID paper's IndexTTS results stay reproducible.
⚠️ Licensing — read before use. Upstream
IndexTeam/IndexTTS-1.5is taggedapache-2.0on the Hub, but the bundledindex-tts-src/LICENSEis the bilibili Model Use License Agreement, which carries its own restrictions. These are not the same terms. Review both, and defer to upstream, before any use beyond research.
| Upstream | IndexTeam/IndexTTS-1.5 |
| Pinned commit | 25851a6036dfd3095bb70fb3c8f49217104672c3 |
| Modifications | None to the weights. Adds a vendored index-tts-src/ tree. |
| Purpose | Reproducibility anchor for the cond-ID paper |
This is the base checkpoint that cond-ID edits — not an unlearned model.
gpt.pth GPT backbone (1.17 GB)
bigvgan_generator.pth BigVGAN vocoder generator (536 MB)
bigvgan_discriminator.pth BigVGAN discriminator (1.65 GB, training only —
not needed for inference)
dvae.pth discrete VAE (243 MB)
bpe.model tokenizer
config.yaml model config
index-tts-src/ vendored upstream index-tts source tree
bigvgan_discriminator.pth is mirrored only for completeness with upstream; inference
never loads it. Skip it to save ~1.65 GB:
from huggingface_hub import snapshot_download
snapshot_download(
repo_id="RootAccess4Life/cond-id-indextts-1.5",
local_dir="models/index_tts",
ignore_patterns=["bigvgan_discriminator.pth", "index-tts-src/*"],
)
Then:
export CD_INDEXTTS_DIR=models/index_tts
python scripts/run_eval.py --config configs/indextts.yaml
IndexTTS needs its own isolated environment; the repo automates that with
bash scripts/setup_indextts.sh.
IndexTTS-1.5 is architecturally the most distinct of the three backbones: autoregressive but perceiver-conditioned, with a BigVGAN vocoder. It is the strongest test of whether cond-ID is genuinely host-independent rather than tuned to one conditioning design.
Erasure on IndexTTS is partial. BigVGAN carries its own speaker encoder, which is a second identity path that a conditioning-path edit does not touch. cond-ID erases what flows through the conditioning path; residual identity can still leak through the vocoder. This is reported as a limitation in the paper rather than smoothed over.
@misc{pujari_condid,
title = {cond-ID: Conditioning-Space Identity Redirection for Speaker Unlearning in Zero-Shot TTS},
author = {Pujari, Aditya and Rattani, Ajita},
note = {Preprint},
}
Please also cite IndexTTS upstream. All credit for these weights belongs to IndexTeam.