Downloads · 30 days
12
6% of all-time downloads
alankessler/Mistral-Nemo-Instruct-2407-MLX-2bit
Mistral-Nemo-Instruct-2407-MLX-2bit is a machine learning model from alankessler. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for mlx. The card lists the license as apache-2.0.
MLX quantized version of Mistral Nemo Instruct 2407.
Downloads · 30 days
12
6% of all-time downloads
All-time downloads
186
Public
Parameters
12.2B
3.8 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors3.8 GB · 100%
How the weights are stored.
U3212.2B · 100%
From the Hugging Face model README
MLX quantized version of Mistral Nemo Instruct 2407.
pip install mlx-lm
from mlx_lm import load, generate
model, tokenizer = load("alankessler/Mistral-Nemo-Instruct-2407-MLX-2bit")
prompt = tokenizer.apply_chat_template(
[{"role": "user", "content": "Hello!"}],
add_generation_prompt=True,
tokenize=False,
)
response = generate(model, tokenizer, prompt=prompt, max_tokens=512)
print(response)