Downloads · 30 days
527
100% of all-time downloads
Akahsizrr/Qwen3.8-27B-Code-Tools-Merged
Qwen3.8-27B-Code-Tools-Merged is a text generation model from Akahsizrr. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
Full merged checkpoint of the Qwen3.8-27B coding/reasoning/tool-calling fine-tune stack.
Downloads · 30 days
527
100% of all-time downloads
All-time downloads
527
Public
Parameters
26.9B
53.8 GB on disk
Likes
3
Public
Click a slice to open those files.
.safetensors53.8 GB · 100%
From the Hugging Face model README
Full merged checkpoint of the Qwen3.8-27B coding/reasoning/tool-calling fine-tune stack.
Qwen/Qwen3.8-27B (base)Akahsizrr/qwen3.8-27b-lora-xhigh-code-tools-1k (LoRA, merged)Akahsizrr/qwen3.8-27b-lora-kimi-k3-tools-code-instr (LoRA, merged)Akahsizrr/qwen3.8-27b-lora-comp-v1 (LoRA, merged)LiveCodeBench v6 (test6.jsonl, 175 problems, pass@1, n=1, temp 0.2, top_p 0.95,
max_tokens 32768, xhigh reasoning effort, vLLM 0.28.0):
from vllm import LLM, SamplingParams
llm = LLM(model="Akahsizrr/Qwen3.8-27B-Code-Tools-Merged", dtype="bfloat16",
max_model_len=36864, gpu_memory_utilization=0.90, enable_prefix_caching=True)
sp = SamplingParams(n=1, max_tokens=32768, temperature=0.2, top_p=0.95)
out = llm.chat(messages, sp, chat_template_kwargs={"reasoning_effort": "xhigh"})
Reasoning is emitted between <tool_call>/</think> ; extract the final code block from the answer.