Downloads · 30 days
254
17% of all-time downloads
jaweed123/Qwen3.5-0.8B-Python-SFT
Qwen3.5-0.8B-Python-SFT is a text generation model from jaweed123. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
Python code generation model — Qwen3.5-0.8B-Base fine-tuned with QLoRA (Supervised Fine-Tuning) on CodeSearchNet (Python): docstring → function code pairs.
Downloads · 30 days
254
17% of all-time downloads
All-time downloads
1.5K
Public
Parameters
752M
7.3 GB on disk
Likes
0
Public
Click a slice to open those files.
.gguf2.9 GB · 65%
How the weights are stored.
BF16752M · 100%
From the Hugging Face model README
Python code generation model — Qwen3.5-0.8B-Base fine-tuned with QLoRA (Supervised Fine-Tuning) on CodeSearchNet (Python): docstring → function code pairs.
| Property | Value |
|---|---|
| Base model | Qwen/Qwen3.5-0.8B-Base |
| Method | QLoRA (4-bit base + LoRA r=16, alpha=32) |
| Trainable params | 6.4M / 759M (0.84%) |
| Dataset | CodeSearchNet Python — 408K samples (13,590 repos) |
| Task | Docstring → Python function code |
| Sequence length | 2048 |
| Precision | BF16 |
| Hardware | NVIDIA RTX 4060 8GB |
| Metric | Value |
|---|---|
| Train loss | 0.330 |
| Eval loss | 1.214 |
| Steps | 25,524 (1 epoch) |
| Runtime | ~25.7h |
pass@1 (temperature 0.2), official test harness, both models in bf16.
| Benchmark | Base | Fine-tuned | Improvement |
|---|---|---|---|
| HumanEval | 1.2% | 17.7% | 14.5x |
| MBPP | 0.0% | 0.2% | 0 → 1 |
Full report with example solutions: reports/evaluation_report.md in the training repo.
from unsloth import FastLanguageModel
model, tokenizer = FastLanguageModel.from_pretrained(
"Qwen/Qwen3.5-0.8B-Base",
max_seq_length=2048,
load_in_4bit=True,
)
model, tokenizer = FastLanguageModel.from_pretrained(
"jaweed123/Qwen3.5-0.8B-Python-SFT",
max_seq_length=2048,
load_in_4bit=True,
)
# llama.cpp
llama-cli -m qwen3.5-0.8b-python-sft-q4_k_m.gguf -p "Write a Python function that..."
# Ollama
ollama create qwen3.5-python -f Modelfile
# Modelfile
FROM qwen3.5-0.8b-python-sft-q4_k_m.gguf
TEMPLATE "{{ if .System }}<|im_start|>system
{{ .System }}<|im_end|>
{{ end }}<|im_start|>user
{{ .Prompt }}<|im_end|>
<|im_start|>assistant
"
| File | Description |
|---|---|
adapter_model.safetensors | LoRA adapter (small, ~13MB) |
model.safetensors | Merged 16-bit model |
qwen3.5-0.8b-python-sft-q4_k_m.gguf | GGUF Q4_K_M (~0.5GB) |
qwen3.5-0.8b-python-sft-q8_0.gguf | GGUF Q8_0 (~0.9GB) |
qwen3.5-0.8b-python-sft-f16.gguf | GGUF F16 |