Downloads · 30 days
51
45% of all-time downloads
ChrisColeTech/scenema-audio
scenema-audio is a text-to-audio model from ChrisColeTech. Use it for the text-to-audio task on the model card, and read the license before you ship it in a product. The card lists the license as other.
Downloads · 30 days
51
45% of all-time downloads
All-time downloads
113
Public
Repo size
21.3 GB
Likes
2
Public
Click a slice to open those files.
.safetensors13.1 GB · 61%
From the Hugging Face model README
<audio controls src="https://huggingface.co/ChrisColeTech/scenema-audio/resolve/main/samples/ComfyUI_00024.mp3"></audio>
<audio controls src="https://huggingface.co/ChrisColeTech/scenema-audio/resolve/main/samples/ComfyUI_00030.mp3"></audio>
| Component | File | ~Size | Download |
|---|---|---|---|
| Transformer | scenema-audio-transformer-int8.safetensors | 4.91 GB | Link |
| Text Encoder | gemma-3-12b-it-Q4_K_M.gguf | 7.3 GB | Link |
| Audio Pipeline VAE | scenema-audio-pipeline.safetensors | 6.71 GB | Link |
| Audio Encoder VAE | scenema-audio-vae-encoder.safetensors | 42.7 MB | Link |
| Extras folder | scenema-audio/extras | 2.3 GB | Link |
⚠ scenema-audio extras is required for longer audio, and for voice-to-voice ⚠
scenema-audio/extras folder if it doesnt existMust have:
In comfy:

Use the workflow Link

| Base model | ScenemaAI/scenema-audio |