Downloads · 30 days
53
10% of all-time downloads
ssdataanalysis/DictaLM-3.0-Nemotron-12B-Instruct-mlx-8Bit
DictaLM-3.0-Nemotron-12B-Instruct-mlx-8Bit is a text generation model from ssdataanalysis. Use it when you need the model to write or continue text. It is set up for mlx. The card lists the license as other.
The Model ssdataanalysis/DictaLM-3.0-Nemotron-12B-Instruct-mlx-8Bit was converted to MLX format from dicta-il/DictaLM-3.0-Nemotron-12B-Instruct using mlx-lm version 0.29.1.
Downloads · 30 days
53
10% of all-time downloads
All-time downloads
538
Public
Parameters
3.5B
13.1 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors13.1 GB · 100%
How the weights are stored.
U323.1B · 89%
From the Hugging Face model README
The Model ssdataanalysis/DictaLM-3.0-Nemotron-12B-Instruct-mlx-8Bit was converted to MLX format from dicta-il/DictaLM-3.0-Nemotron-12B-Instruct using mlx-lm version 0.29.1.
pip install mlx-lm
from mlx_lm import load, generate
model, tokenizer = load("ssdataanalysis/DictaLM-3.0-Nemotron-12B-Instruct-mlx-8Bit")
prompt="hello"
if hasattr(tokenizer, "apply_chat_template") and tokenizer.chat_template is not None:
messages = [{"role": "user", "content": prompt}]
prompt = tokenizer.apply_chat_template(
messages, tokenize=False, add_generation_prompt=True
)
response = generate(model, tokenizer, prompt=prompt, verbose=True)