Downloads · 30 days
565
54% of all-time downloads
HanzoHuang/Qwen2.5-3B-Instruct-RKLLM
Qwen2.5-3B-Instruct-RKLLM is a text generation model from HanzoHuang. Use it when you need the model to write or continue text. It is set up for rkllm. The card lists the license as apache-2.0.
RKLLM-converted Qwen2.5-3B-Instruct language-model artifacts for Rockchip RK3576 and RK3588 NPUs.
Downloads · 30 days
565
54% of all-time downloads
All-time downloads
1.1K
Public
Repo size
9.9 GB
Likes
0
Public
Click a slice to open those files.
.rkllm9.9 GB · 100%
From the Hugging Face model README
RKLLM-converted Qwen2.5-3B-Instruct language-model artifacts for Rockchip RK3576 and RK3588 NPUs.
These hardware-specific .rkllm files require a compatible Rockchip RKLLM runtime. They are not Transformers checkpoints and cannot be loaded directly with Transformers, llama.cpp, or Ollama.
Review the upstream license and usage restrictions before use or redistribution.
RKLLM Toolkit: v1.2.3
Use a file built for the exact target SoC.
| Target | Quantization | File | SHA256 |
|---|---|---|---|
| RK3576 | W4A16 | Qwen2.5-3B-Instruct_RK3576_w4a16.rkllm | 5f2480e10a794848c8d4a5a21a61c015d96570f2a58e4543897b694b59576908 |
| RK3576 | W8A8 | Qwen2.5-3B-Instruct_RK3576_w8a8.rkllm | 5a14ed85f65d3c2890ef8e2b4ab9f094bf9b0992ddebd6f2b9ad8d3b539897b9 |
| RK3588 | W8A8 | Qwen2.5-3B-Instruct_RK3588_w8a8.rkllm | 054a4ac54ea7d483ac17431df5286eb6b2a81d351fbb44b55dc4491f71a7ea46 |
The repository also includes Qwen2.5-3B-Instruct_data_quant.json, used as calibration data during conversion.
hf download HanzoHuang/Qwen2.5-3B-Instruct-RKLLM \
RK3576/Qwen2.5-3B-Instruct_RK3576_w4a16.rkllm \
--local-dir Qwen2.5-3B-Instruct-RKLLM
Use the Qwen2.5 Instruct chat template with the RKLLM runtime. For Docker deployment, see Hanzo-Huang/rkllm-docker.
These are target-specific converted artifacts. Validate quality and runtime compatibility on your Rockchip device.
Thanks to the Qwen Team, Rockchip, and the RKLLM community.