Downloads · 30 days
13
10% of all-time downloads
moot20/DeepScaleR-1.5B-Preview-MLX-8bits
DeepScaleR-1.5B-Preview-MLX-8bits is a text generation model from moot20. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as mit.
The Model moot20/DeepScaleR-1.5B-Preview-MLX-8bits was converted to MLX format from agentica-org/DeepScaleR-1.5B-Preview using mlx-lm version 0.21.1.
Downloads · 30 days
13
10% of all-time downloads
All-time downloads
132
Public
Parameters
1.8B
1.9 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors1.9 GB · 99%
How the weights are stored.
U321.8B · 100%
From the Hugging Face model README
The Model moot20/DeepScaleR-1.5B-Preview-MLX-8bits was converted to MLX format from agentica-org/DeepScaleR-1.5B-Preview using mlx-lm version 0.21.1.
pip install mlx-lm
from mlx_lm import load, generate
model, tokenizer = load("moot20/DeepScaleR-1.5B-Preview-MLX-8bits")
prompt = "hello"
if tokenizer.chat_template is not None:
messages = [{"role": "user", "content": prompt}]
prompt = tokenizer.apply_chat_template(
messages, add_generation_prompt=True
)
response = generate(model, tokenizer, prompt=prompt, verbose=True)