Skip to content

Lwasinam

voicera

Lwasinam/voicera

voicera is a text-to-speech model from Lwasinam. Use it when you need text read aloud. It is set up for transformers.

Voicera is a AR text-to-speech model trained on ~1000hrs of speech data. speech is converted to discrete tokens using "Multi-Scale Neural Audio Codec (SNAC)" model NB: This is not a SOTA model, and not accuarate enoug…

Downloads · 30 days

238

21% of all-time downloads

All-time downloads

1.1K

Public

Parameters

304M

1.2 GB on disk

Likes

24

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors1.2 GB · 100%

At a glance

Task
Text-to-Speech
Library
transformers
Model type
gpt2
Access
Public
Created
Jul 26, 2024
Updated
Jul 30, 2024
SHA
f74ea70b
Task
Text-to-Speech
Library
transformers
Type
gpt2
Created
Jul 26, 2024
Updated
Jul 30, 2024