Downloads · 30 days
20
4% of all-time downloads
alankessler/Mistral-Nemo-Instruct-2407-MLX-3bit
Mistral-Nemo-Instruct-2407-MLX-3bit is a machine learning model from alankessler. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for mlx. The card lists the license as apache-2.0.
MLX quantized version of Mistral Nemo Instruct 2407.
Downloads · 30 days
20
4% of all-time downloads
All-time downloads
476
Public
Parameters
12.2B
5.4 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors5.4 GB · 100%
How the weights are stored.
U3212.2B · 100%
From the Hugging Face model README
MLX quantized version of Mistral Nemo Instruct 2407.
pip install mlx-lm
from mlx_lm import load, generate
model, tokenizer = load("alankessler/Mistral-Nemo-Instruct-2407-MLX-3bit")
prompt = tokenizer.apply_chat_template(
[{"role": "user", "content": "Hello!"}],
add_generation_prompt=True,
tokenize=False,
)
response = generate(model, tokenizer, prompt=prompt, max_tokens=512)
print(response)