Downloads · 30 days
327
42% of all-time downloads
HanzoHuang/Qwen2.5-1.5B-Instruct-RKLLM
Qwen2.5-1.5B-Instruct-RKLLM is a text generation model from HanzoHuang. Use it when you need the model to write or continue text. It is set up for rkllm. The card lists the license as apache-2.0.
RKLLM-converted Qwen2.5-1.5B-Instruct language-model artifacts for Rockchip RK3576 and RK3588 NPUs.
Downloads · 30 days
327
42% of all-time downloads
All-time downloads
770
Public
Repo size
5.5 GB
Likes
0
Public
Click a slice to open those files.
.rkllm5.5 GB · 100%
From the Hugging Face model README
RKLLM-converted Qwen2.5-1.5B-Instruct language-model artifacts for Rockchip RK3576 and RK3588 NPUs.
These hardware-specific .rkllm files require a compatible Rockchip RKLLM runtime. They are not Transformers checkpoints and cannot be loaded directly with Transformers, llama.cpp, or Ollama.
RKLLM Toolkit: v1.2.3
Use a file built for the exact target SoC.
| Target | Quantization | File | SHA256 |
|---|---|---|---|
| RK3576 | W4A16 | Qwen2.5-1.5B-Instruct_RK3576_w4a16.rkllm | 70b4b3c54221892cc403aa877612e0bfdc24abeda5e0011ee3a2e9a3c24e332c |
| RK3576 | W8A8 | Qwen2.5-1.5B-Instruct_RK3576_w8a8.rkllm | a2ab655ca7e6a3d3626a9a93e69a3a95660cedd7d4820ddf041371b45f4a6c94 |
| RK3588 | W8A8 | Qwen2.5-1.5B-Instruct_RK3588_w8a8.rkllm | e296d41ff1ff64f86e293227f15e35520dbeacebb74d21592a0790919691ca2d |
The repository also includes Qwen2.5-1.5B-Instruct_data_quant.json, used as calibration data during conversion.
hf download HanzoHuang/Qwen2.5-1.5B-Instruct-RKLLM \
RK3576/Qwen2.5-1.5B-Instruct_RK3576_w4a16.rkllm \
--local-dir Qwen2.5-1.5B-Instruct-RKLLM
Use the Qwen2.5 Instruct chat template with the RKLLM runtime. For Docker deployment, see Hanzo-Huang/rkllm-docker.
These are target-specific converted artifacts. Validate quality and runtime compatibility on your Rockchip device.
Thanks to the Qwen Team, Rockchip, and the RKLLM community.