Downloads · 30 days
0
mobiletransformers/Qwen2.5-0.5B-Instruct
Qwen2.5-0.5B-Instruct is a text generation model from mobiletransformers. Use it when you need the model to write or continue text. It is set up for mobiletransformers.
Downloads · 30 days
0
Access
Public
Updated Aug 17, 2026
Repo size
3.2 GB
Likes
0
Public
Click a slice to open those files.
.data2.4 GB · 76%
From the Hugging Face model README

On-device (Android) package exported from Qwen/Qwen2.5-0.5B-Instruct with MobileTransformers.
core — shared files every other group needsinference — generate or score on devicetrain — fine-tune on device, then merge the adapter back into the base weightsrag — retrieve over documents you ingest, and ground answers in them8q_proj, v_projQwen/Qwen2.5-0.5B-Instructtext-generation-with-past| id | EP | quant | engines | features | min API | rec. RAM (MB) |
|---|---|---|---|---|---|---|
| cpu-int4 | cpu | int4 | native | core, inference, train, rag | 28 | — |
Default variant: cpu-int4.
This is a MobileTransformers package, not a plain Hugging Face model: it is a manifest plus per-variant ONNX stages and a weight-handoff map. transformers, optimum and plain onnxruntime cannot load it. Use the framework:
https://github.com/martinkorelic/mobiletransformers
// Android — pulls, verifies and installs on first use.
val model = MobileTransformers.fromPretrained(
context = context,
repoId = "mobiletransformers/Qwen2.5-0.5B-Instruct",
)
# Host — download and inspect the package without a device.
mobiletransformers pull --repo-id mobiletransformers/Qwen2.5-0.5B-Instruct
If you are using this framework for your own work, please cite:
@misc{mobiletransformers2025,
author = {Koreli\v{c}, Martin and Pejovi{\'c}, Veljko},
title = {MobileTransformers: An On-Device LLM PEFT Framework for Fine-Tuning and Inference},
year = {2025},
howpublished = {\url{https://gitlab.fri.uni-lj.si/lrk/mobiletransformers}}
}