Downloads · 30 days
902
77% of all-time downloads
blackshell69/Qwen3.5-9B-distilled-coding-agentic
Qwen3.5-9B-distilled-coding-agentic is a text generation model from blackshell69. Use it when you need the model to write or continue text. It is set up for peft. The card lists the license as apache-2.0.
LoRA adapter for Qwen/Qwen3.5-9B, trained via sequence-level knowledge distillation from Qwen3.8-27B (teacher, served locally via llama.cpp) to strengthen coding and agentic (tool-calling) capability in a model small…
Downloads · 30 days
902
77% of all-time downloads
All-time downloads
1.2K
Public
Repo size
6 GB
Likes
4
Trending 1
Click a slice to open those files.
.gguf5.7 GB · 96%
From the Hugging Face model README
LoRA adapter for Qwen/Qwen3.5-9B, trained via
sequence-level knowledge distillation from Qwen3.8-27B (teacher, served locally via
llama.cpp) to strengthen coding and agentic (tool-calling) capability in a model small
enough to run on consumer GPUs.
q/k/v/o_proj, gate/up/down_proj), bf16 compute.ise-uiuc/Magicoder-OSS-Instruct-75K;
reference solutions discarded, the teacher writes its own.glaiveai/glaive-function-calling-v2;
every assistant turn (tool calls and post-tool-response answers alike) is regenerated
by the teacher conditioned on the recorded history, in OpenAI tools/tool_calls format.from transformers import AutoModelForCausalLM, AutoTokenizer
from peft import PeftModel
base = AutoModelForCausalLM.from_pretrained("Qwen/Qwen3.5-9B", dtype="bfloat16")
model = PeftModel.from_pretrained(base, "blackshell69/Qwen3.5-9B-distilled-coding-agentic")
tokenizer = AutoTokenizer.from_pretrained("Qwen/Qwen3.5-9B")
Two ways to run this with llama.cpp:
Single merged file (Qwen3.5-9B-distilled-IQ4_NL.gguf, recommended):
llama-server -m Qwen3.5-9B-distilled-IQ4_NL.gguf --jinja
Base + LoRA adapter (Qwen3.5-9B-distilled-LoRA-F16.gguf), if you'd rather keep the
adapter separate from a base GGUF you already have — numerically equivalent to the merged
file:
llama-server -m Qwen3.5-9B-<any-quant>.gguf \
--lora Qwen3.5-9B-distilled-LoRA-F16.gguf --jinja