Downloads · 30 days
13
11% of all-time downloads
Indus-Labs/v2_saavi_devi_snor
v2_saavi_devi_snor is a text generation model from Indus-Labs. Use it when you need the model to write or continue text. The card lists the license as llama3.2.
This is a fine-tuned version of snorbyte/snorTTS-Indic-v0 specialized for Hinglish (Hindi-English mixed) text-to-speech generation.
Downloads · 30 days
13
11% of all-time downloads
All-time downloads
118
Public
Parameters
3.3B
6.6 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors6.6 GB · 100%
From the Hugging Face model README
This is a fine-tuned version of snorbyte/snorTTS-Indic-v0 specialized for Hinglish (Hindi-English mixed) text-to-speech generation.
from transformers import AutoTokenizer, AutoModelForCausalLM
import torch
# Load model and tokenizer
model_name = "Indus-Labs/v2_saavi_devi_snor"
tokenizer = AutoTokenizer.from_pretrained(model_name)
model = AutoModelForCausalLM.from_pretrained(
model_name,
torch_dtype=torch.float16,
device_map="auto"
)
# Generate text
prompt = "Hello doston, main aapka dost hun"
inputs = tokenizer(prompt, return_tensors="pt")
outputs = model.generate(**inputs, max_new_tokens=1200)
This model generates audio tokens that need to be decoded using a SNAC (Scalable Neural Audio Codec) model:
from snac import SNAC
# Load SNAC decoder
snac_model = SNAC.from_pretrained("hubertsiuzdak/snac_24khz")
# Process generated tokens to audio codes and decode
# (See full implementation in the original training code)
If you use this model, please cite the original base model:
@misc{canopylabs-3b-hi,
title={3B Hindi Pretrained Model},
author={Canopy Labs},
year={2024},
url={https://huggingface.co/snorbyte/snorTTS-Indic-v0}
}