Downloads · 30 days
96
3% of all-time downloads
BlueMoonlight/DeepSeek-R1-Distill-Qwen-14B-mlx-5Bit
DeepSeek-R1-Distill-Qwen-14B-mlx-5Bit is a text generation model from BlueMoonlight. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as mit.
The Model BlueMoonlight/DeepSeek-R1-Distill-Qwen-14B-mlx-5Bit was converted to MLX format from deepseek-ai/DeepSeek-R1-Distill-Qwen-14B using mlx-lm version 0.29.1.
Downloads · 30 days
96
3% of all-time downloads
All-time downloads
2.9K
Public
Parameters
14.8B
10.2 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors10.2 GB · 100%
How the weights are stored.
U3214.8B · 100%
From the Hugging Face model README
The Model BlueMoonlight/DeepSeek-R1-Distill-Qwen-14B-mlx-5Bit was converted to MLX format from deepseek-ai/DeepSeek-R1-Distill-Qwen-14B using mlx-lm version 0.29.1.
pip install mlx-lm
from mlx_lm import load, generate
model, tokenizer = load("BlueMoonlight/DeepSeek-R1-Distill-Qwen-14B-mlx-5Bit")
prompt="hello"
if hasattr(tokenizer, "apply_chat_template") and tokenizer.chat_template is not None:
messages = [{"role": "user", "content": prompt}]
prompt = tokenizer.apply_chat_template(
messages, tokenize=False, add_generation_prompt=True
)
response = generate(model, tokenizer, prompt=prompt, verbose=True)