Downloads · 30 days
40
59% of all-time downloads
GatekeeperZA/Qwen3.5-0.8B-RKLLM-v1.2.3
Qwen3.5-0.8B-RKLLM-v1.2.3 is a machine learning model from GatekeeperZA. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for rkllm. The card lists the license as apache-2.0.
RKLLM conversion of Qwen/Qwen3.5-0.8B for Rockchip RK3588 NPU inference.
Downloads · 30 days
40
59% of all-time downloads
All-time downloads
68
Public
Repo size
1.3 GB
Likes
0
Public
Click a slice to open those files.
.rkllm1.3 GB · 100%
From the Hugging Face model README
RKLLM conversion of Qwen/Qwen3.5-0.8B for Rockchip RK3588 NPU inference.
Converted with RKLLM Toolkit v1.2.3. This is the base model variant — suitable for custom fine-tunes or applications that supply their own system prompting.
| Property | Value |
|---|---|
| Base Model | Qwen/Qwen3.5-0.8B |
| Toolkit Version | RKLLM Toolkit v1.2.3 |
| Runtime Version | RKLLM Runtime ≥ v1.2.1 (v1.2.3 recommended) |
| Quantization | w8a8 (8-bit weights, 8-bit activations) |
| Target Platform | RK3588 |
| NPU Cores | 3 |
| Thinking Mode | ❌ Not applicable (base model) |
| Languages | English, Chinese (multilingual) |
Qwen3.5-0.8B is the successor to Qwen3-0.6B, offering improved architecture and training. At under 1GB after quantisation it is the smallest model in the RK3588 lineup — ideal for low-latency responses, always-on assistants, or edge deployments where memory is tight.
mkdir -p ~/models/Qwen3.5-0.8B
cd ~/models/Qwen3.5-0.8B
git lfs install && git clone https://huggingface.co/GatekeeperZA/Qwen3.5-0.8B-RKLLM-v1.2.3 .
Use with GatekeeperZA/RKLLM-API-Server.
git clone https://github.com/airockchip/rknn-llm.git
cd rknn-llm/examples/rkllm_api_demo
./build/rkllm_api_demo /path/to/Qwen3.5-0.8B-rk3588-w8a8.rkllm 4096 8192
| File | Description |
|---|---|
Qwen3.5-0.8B-rk3588-w8a8.rkllm | Quantized model for RK3588 NPU |