Downloads · 30 days
404
46% of all-time downloads
mlx-community/MiniMax-Music3-4bit
MiniMax-Music3-4bit is a text-to-audio model from mlx-community. Use it for the text-to-audio task on the model card, and read the license before you ship it in a product. It is set up for mlx-audio. The card lists the license as other.
Native MLX affine 4-bit weights for MiniMaxAI/MiniMax-Music3, converted for lyric-conditioned song generation on Apple Silicon with mlx-audio.
Downloads · 30 days
404
46% of all-time downloads
All-time downloads
873
Public
Parameters
11.7B
9.2 GB on disk
Likes
7
Trending 1
Click a slice to open those files.
.safetensors9.2 GB · 100%
How the weights are stored.
U329.9B · 85%
From the Hugging Face model README
Native MLX affine 4-bit weights for
MiniMaxAI/MiniMax-Music3,
converted for lyric-conditioned song generation on Apple Silicon with
mlx-audio.
Community conversion, not an official MiniMax release. All model credit goes to MiniMax. Review the original model card and license before use.
Other MLX variants:
BF16
· 8-bit
· 6-bit
· MXFP8
· MXFP4 (experimental)
· NVFP4 (experimental)
MiniMax Music 3 support was merged upstream in
Blaizzy/mlx-audio#888.
Until a PyPI release includes it, install the upstream merge commit directly:
python -m pip install "mlx-audio @ git+https://github.com/Blaizzy/mlx-audio.git@784b29e2691a93ca7483147d86f61859dfaa6296"
python -m mlx_audio.music.generate \
--model mlx-community/MiniMax-Music3-4bit \
--caption "Warm acoustic pop, 96 BPM, intimate female vocal" \
--lyrics $'[verse]\nMorning light across the room\n[chorus]\nSing with me' \
--duration 30 \
--steps 30 \
--seed 7 \
--output song.wav
from mlx_audio.music import load
model = load("mlx-community/MiniMax-Music3-4bit")
result = next(
model.generate(
text="Warm acoustic pop, 96 BPM, intimate female vocal",
lyrics="[verse]\nMorning light across the room\n[chorus]\nSing with me",
duration=30,
steps=30,
seed=7,
)
)
print(result.audio.shape, result.sample_rate) # stereo, 44100 Hz
Lyrics are required by the checkpoint contract. Use [instrumental] explicitly
for instrumental generation. Duration is a requested upper bound: the
autoregressive stage may emit its end token early. Style, tempo, instrument, and
vocal controls are probabilistic rather than strict.
Converted with mlx-audio 0.4.8 development commit c2fa486 and MLX 0.31.2.
The weights remain subject to the
MiniMax-Music3 Community License,
including its acceptable-use and commercial terms. The full license text is
included in this repository.