Downloads · 30 days
134
15% of all-time downloads
mlx-community/MiniMax-Music3-mxfp4
MiniMax-Music3-mxfp4 is a text-to-audio model from mlx-community. Use it for the text-to-audio task on the model card, and read the license before you ship it in a product. It is set up for mlx-audio. The card lists the license as other.
Native MLX MXFP4 weights for MiniMaxAI/MiniMax-Music3, converted for lyric-conditioned song generation on Apple Silicon with mlx-audio. This is an experimental memory-saving variant; use MXFP8 when lyric fidelity matt…
Downloads · 30 days
134
15% of all-time downloads
All-time downloads
896
Public
Parameters
11.7B
8.9 GB on disk
Likes
5
Public
Click a slice to open those files.
.safetensors8.9 GB · 100%
How the weights are stored.
U329.9B · 85%
From the Hugging Face model README
Native MLX MXFP4 weights for
MiniMaxAI/MiniMax-Music3,
converted for lyric-conditioned song generation on Apple Silicon with
mlx-audio. This is an experimental
memory-saving variant; use MXFP8 when lyric fidelity matters.
Community conversion, not an official MiniMax release. All model credit goes to MiniMax. Review the original model card and license before use.
Other MLX variants:
BF16
· 8-bit
· 6-bit
· 4-bit
· MXFP8 (recommended)
· NVFP4 (experimental)
MiniMax Music 3 support was merged upstream in
Blaizzy/mlx-audio#888.
Until a PyPI release includes it, install the upstream merge commit directly:
python -m pip install "mlx-audio @ git+https://github.com/Blaizzy/mlx-audio.git@784b29e2691a93ca7483147d86f61859dfaa6296"
python -m mlx_audio.music.generate \
--model mlx-community/MiniMax-Music3-mxfp4 \
--caption "Warm acoustic pop, 96 BPM, intimate female vocal" \
--lyrics $'[verse]\nMorning light across the room\n[chorus]\nSing with me' \
--duration 30 \
--steps 30 \
--seed 7 \
--output song.wav
from mlx_audio.music import load
model = load("mlx-community/MiniMax-Music3-mxfp4")
result = next(
model.generate(
text="Warm acoustic pop, 96 BPM, intimate female vocal",
lyrics="[verse]\nMorning light across the room\n[chorus]\nSing with me",
duration=30,
steps=30,
seed=7,
)
)
print(result.audio.shape, result.sample_rate) # stereo, 44100 Hz
Lyrics are required by the checkpoint contract. Use [instrumental] explicitly
for instrumental generation. Duration is a requested upper bound: the
autoregressive stage may emit its end token early. Style, tempo, instrument, and
vocal controls are probabilistic rather than strict.
Converted with mlx-audio 0.4.8 development commit c2fa486 and MLX 0.31.2.
The weights remain subject to the
MiniMax-Music3 Community License,
including its acceptable-use and commercial terms. The full license text is
included in this repository.