Skip to content

zhangj1an

AudioX

zhangj1an/AudioX

AudioX is a text-to-audio model from zhangj1an. Use it for the text-to-audio task on the model card, and read the license before you ship it in a product. It is set up for diffusers. The card lists the license as cc-by-nc-4.0.

AudioX is a unified framework for generating audio and music from diverse multimodal control signals, including text, video, and audio. It features a Multimodal Adaptive Fusion (MAF) module to effectively align and fu…

Downloads · 30 days

2.1K

65% of all-time downloads

All-time downloads

3.3K

Public

Parameters

2.8B

76.4 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors11.7 GB · 100%

Parameter types

How the weights are stored.

F322.7B · 97%

Base models

Task
Text-to-Audio
Library
diffusers
Type
diffusion_cond
License
cc-by-nc-4.0
Created
Apr 6, 2026
Updated
Apr 17, 2026
AudioX — AI Model — AIMarketly