Downloads · 30 days
89
18% of all-time downloads
mlx-community/MiniMax-Music3-nvfp4
MiniMax-Music3-nvfp4 is a text-to-audio model from mlx-community. Use it for the text-to-audio task on the model card, and read the license before you ship it in a product. It is set up for mlx-audio. The card lists the license as other.
Native MLX NVFP4 weights for MiniMaxAI/MiniMax-Music3, converted for lyric-conditioned song generation on Apple Silicon with mlx-audio. This variant is experimental until it receives broader listening evaluation.
Downloads · 30 days
89
18% of all-time downloads
All-time downloads
503
Public
Parameters
11.7B
9.2 GB on disk
Likes
2
Public
Click a slice to open those files.
.safetensors9.2 GB · 100%
How the weights are stored.
U329.9B · 85%
From the Hugging Face model README
Native MLX NVFP4 weights for
MiniMaxAI/MiniMax-Music3,
converted for lyric-conditioned song generation on Apple Silicon with
mlx-audio. This variant is experimental
until it receives broader listening evaluation.
Community conversion, not an official MiniMax release. All model credit goes to MiniMax. Review the original model card and license before use.
Other MLX variants:
BF16
· 8-bit
· 6-bit
· 4-bit
· MXFP8
· MXFP4 (experimental)
MiniMax Music 3 support was merged upstream in
Blaizzy/mlx-audio#888.
Until a PyPI release includes it, install the upstream merge commit directly:
python -m pip install "mlx-audio @ git+https://github.com/Blaizzy/mlx-audio.git@784b29e2691a93ca7483147d86f61859dfaa6296"
python -m mlx_audio.music.generate \
--model mlx-community/MiniMax-Music3-nvfp4 \
--caption "Warm acoustic pop, 96 BPM, intimate female vocal" \
--lyrics $'[verse]\nMorning light across the room\n[chorus]\nSing with me' \
--duration 30 \
--steps 30 \
--seed 7 \
--output song.wav
from mlx_audio.music import load
model = load("mlx-community/MiniMax-Music3-nvfp4")
result = next(
model.generate(
text="Warm acoustic pop, 96 BPM, intimate female vocal",
lyrics="[verse]\nMorning light across the room\n[chorus]\nSing with me",
duration=30,
steps=30,
seed=7,
)
)
print(result.audio.shape, result.sample_rate) # stereo, 44100 Hz
Lyrics are required by the checkpoint contract. Use [instrumental] explicitly
for instrumental generation. Duration is a requested upper bound: the
autoregressive stage may emit its end token early. Style, tempo, instrument, and
vocal controls are probabilistic rather than strict.
Converted with mlx-audio 0.4.8 development commit c2fa486 and MLX 0.31.2.
The weights remain subject to the
MiniMax-Music3 Community License,
including its acceptable-use and commercial terms. The full license text is
included in this repository.