Downloads · 30 days
34
3% of all-time downloads
petergilani/MiniMax-M2.5-mix3-6bit
MiniMax-M2.5-mix3-6bit is a machine learning model from petergilani. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for mlx-lm. The card lists the license as apache-2.0.
--- basemodel: MiniMaxAI/MiniMax-M2.5 language: en libraryname: mlx-lm license: modified-mit modelname: MiniMax-M2.5-mix3-6bit tags: - quantization - mixed36 - minimax - mlx ---
Downloads · 30 days
34
3% of all-time downloads
All-time downloads
986
Public
Parameters
229B
114 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors114 GB · 100%
How the weights are stored.
U32229B · 100%
From the Hugging Face model README
base_model: MiniMaxAI/MiniMax-M2.5 language: en library_name: mlx-lm license: modified-mit model_name: MiniMax-M2.5-mix3-6bit tags:
Mixed precision quantized version of MiniMax M2.5 using mlx-lm with --quant-predicate mixed_3_6.
| Property | Value |
|---|---|
| Base Model | MiniMaxAI/MiniMax-M2.5 |
| Quantization | mlx-lm v0.30.7 with --quant-predicate mixed_3_6 |
| Library | mlx-lm |
| License | modified-mit |
| Parameter | Value |
|---|---|
| temperature | 1.0 |
| top_p | 0.95 |
| top_k | 40 |
import mlx_lm
from mlx_lm.sample_utils import make_sampler
model_path = "petergilani/MiniMax-M2.5-mix3-6bit"
model, tokenizer = mlx_lm.load(model_path)
sampler = make_sampler(temp=1.0, top_p=0.95, top_k=40)
prompt = "Your prompt here"
response = mlx_lm.generate(
model,
tokenizer,
prompt=prompt,
sampler=sampler,
max_tokens=512
)
print(response)