Downloads · 30 days
0
KaaKangSatiry05/rvcintinya
rvcintinya is a machine learning model from KaaKangSatiry05. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
Proyek ini menyediakan pipeline TTS - optional RVC untuk sintesis suara lintas bahasa.
Downloads · 30 days
0
Access
Public
Updated Mar 19, 2026
Repo size
—
Likes
0
Public
Click a slice to open those files.
.py26.5 KB · 81%
From the Hugging Face model README
Proyek ini menyediakan pipeline TTS -> optional RVC untuk sintesis suara lintas bahasa.
Target utamanya:
voice_models.jsonBackend TTS yang tersedia:
xtts: lokal, keluaran WAV, cocok untuk cloning speaker multibahasaedge: cakupan bahasa luas, default paling praktis, keluaran MP3 jika tanpa RVCTahap voice conversion:
rvc-python dipakai sebagai adaptor RVC opsionalpython -m venv .venv
source .venv/bin/activate
pip install -r requirements.txt
Catatan:
xtts butuh model Coqui TTS yang lebih beratedge butuh koneksi internet saat inferensi.pth dan index .index jika tersediaFile registry ada di voice_models.json.
Contoh entri:
[
{
"id": "sample-rvc",
"name": "Sample RVC Voice",
"source": "https://huggingface.co/example-user/example-rvc-model/tree/main"
}
]
source mendukung:
.pth dan opsional .index.pth.zipResolver akan mencari file .pth dan .index otomatis. Jika sumbernya .zip, file akan di-download lalu di-extract ke cache lebih dulu.
Tanpa RVC:
python -m rvc_tts.cli \
--text "Halo, ini uji TTS bahasa Indonesia." \
--lang id \
--backend xtts \
--speaker-wav /path/ke/referensi.wav \
--output ./out.wav
Dengan backend edge:
python -m rvc_tts.cli \
--text "Halo, ini uji TTS multibahasa." \
--lang id-ID \
--backend edge \
--voice id-ID-GadisNeural \
--output ./out.mp3
Dengan RVC:
python -m rvc_tts.cli \
--text "Halo, ini suara hasil TTS lalu masuk RVC." \
--lang id \
--backend xtts \
--speaker-wav /path/ke/referensi.wav \
--rvc-model /path/ke/model.pth \
--rvc-index /path/ke/model.index \
--output ./final.wav
Auto-detect bahasa:
python -m rvc_tts.cli \
--text "こんにちは、音声テストです。" \
--backend edge \
--voice ja-JP-NanamiNeural \
--output ./jp.mp3
Jalankan server:
uvicorn rvc_tts.api:app --host 0.0.0.0 --port 8000
List voice models:
curl http://127.0.0.1:8000/voice-models
Synthesize dengan auto language detect:
curl -X POST http://127.0.0.1:8000/synthesize \
-H "Content-Type: application/json" \
-d '{
"text": "Halo dunia, ini percobaan API.",
"backend": "edge",
"voice": "id-ID-GadisNeural",
"output_format": "mp3"
}'
Synthesize dengan RVC dari registry:
curl -X POST http://127.0.0.1:8000/synthesize \
-H "Content-Type: application/json" \
-d '{
"text": "Halo, ini suara dengan model RVC dari Hugging Face.",
"backend": "xtts",
"speaker_wav": "/path/ke/ref.wav",
"voice_model_id": "sample-rvc",
"output_format": "wav"
}'
Synthesize dengan link Hugging Face ZIP langsung dari user:
curl -X POST http://127.0.0.1:8000/synthesize \
-H "Content-Type: application/json" \
-d '{
"text": "Halo, ini suara dari model zip Hugging Face.",
"backend": "xtts",
"speaker_wav": "/path/ke/ref.wav",
"voice_model_url": "https://huggingface.co/username/repo/resolve/main/model.zip",
"output_format": "wav"
}'
Ambil file audio:
curl -O http://127.0.0.1:8000/audio/<job_id>
/synthesize mengembalikan field expires_at.edge menghasilkan MP3. Jika ingin diproses RVC, gunakan xtts atau tambahkan tahap konversi ke WAV.