Downloads · 30 days
145
19% of all-time downloads
darkmaniac7/TokForge-AccelerationPack-Qwen35-Draft
TokForge-AccelerationPack-Qwen35-Draft is a text generation model from darkmaniac7. Use it when you need the model to write or continue text. The card lists the license as apache-2.0.
- Website: https://tokforge.ai - Discord: https://discord.gg/Acv3CBtfVm - Google Play: https://play.google.com/store/apps/details?id=dev.tokforge - iOS TestFlight: https://testflight.apple.com/join/jnufjzRr
Downloads · 30 days
145
19% of all-time downloads
All-time downloads
744
Public
Repo size
1.4 GB
Likes
0
Public
Click a slice to open those files.
.weight470 MB · 97%
From the Hugging Face model README
Runs on-device in the TokForge app.
Deprecated: this
Qwen3.5-0.8Bdraft bundle is preserved for reproducibility, but the newerQwen3-0.6BTokForge draft line is the practical default.
Two reasons:
Qwen3-0.6B draft line with dense Qwen3 targets measured +34% to +43% faster decode in chat workloads on our test devices.This bundle is preserved because it was an important research step, but it is not the current practical winner. The measured wins live in the same-family Qwen3-0.6B draft line with dense Qwen3 targets; this Qwen3.5 draft lane did not show reliable speedups in our testing.
llm.mnnllm.mnn.weightllm_config.jsonThis repo is for TokForge / MNN users who specifically want to reproduce the older Qwen3.5-0.8B draft path.