Downloads · 30 days
112
29% of all-time downloads
tbhrc/qwen2_5_coder_3b_instruct_4bit
qwen2_5_coder_3b_instruct_4bit is a text generation model from tbhrc. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as other.
The Model mlx-community/Qwen2.5-Coder-3B-Instruct-4bit was converted to MLX format from Qwen/Qwen2.5-Coder-3B-Instruct using mlx-lm version 0.19.3.
Downloads · 30 days
112
29% of all-time downloads
All-time downloads
389
Public
Parameters
3.1B
1.7 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors1.7 GB · 99%
How the weights are stored.
U323.1B · 100%
From the Hugging Face model README
The Model mlx-community/Qwen2.5-Coder-3B-Instruct-4bit was converted to MLX format from Qwen/Qwen2.5-Coder-3B-Instruct using mlx-lm version 0.19.3.
pip install mlx-lm
from mlx_lm import load, generate
model, tokenizer = load("mlx-community/Qwen2.5-Coder-3B-Instruct-4bit")
prompt="hello"
if hasattr(tokenizer, "apply_chat_template") and tokenizer.chat_template is not None:
messages = [{"role": "user", "content": prompt}]
prompt = tokenizer.apply_chat_template(
messages, tokenize=False, add_generation_prompt=True
)
response = generate(model, tokenizer, prompt=prompt, verbose=True)