Downloads Β· 30 days
9
4% of all-time downloads
usernamebetter/nanocoder-v1
nanocoder-v1 is a text generation model from usernamebetter. Use it when you need the model to write or continue text. It is set up for peft. The card lists the license as apache-2.0.
A 4B parameter full-stack coding assistant fine-tuned from Qwen3-4B using Unsloth + LoRA. Trained through a multi-phase pipeline with joint domain training and validation-driven checkpoint selection.
Downloads Β· 30 days
9
4% of all-time downloads
All-time downloads
243
Public
Repo size
2.9 GB
Likes
0
Public
Click a slice to open those files.
.safetensors264 MB Β· 96%
From the Hugging Face model README
A 4B parameter full-stack coding assistant fine-tuned from Qwen3-4B using Unsloth + LoRA. Trained through a multi-phase pipeline with joint domain training and validation-driven checkpoint selection.
Best checkpoint: step 150 β combined score 73.6% across all skill domains.
| Benchmark | Score | Notes |
|---|---|---|
| HumanEval pass@1 | 49.4% | 164 problems, executed against test cases |
| LiveCodeBench | 13.3% | Execution eval on 30 problems (public tests) |
| Frontend (custom) | 58.3% | React, Next.js, TypeScript, Tailwind, a11y |
| Backend (custom) | 87.5% | FastAPI, Express, PostgreSQL, JWT, MongoDB |
| Agent (custom) | 75.0% | Thought β Action β Patch β Reasoning format |
| Combined | 73.6% | Averaged across skill domains |
### Thought β ### Action β ### Patch β ### Reasoning)from unsloth import FastLanguageModel
model, tokenizer = FastLanguageModel.from_pretrained(
model_name="usernamebetter/nanocoder-v1",
max_seq_length=2048,
load_in_4bit=True,
)
FastLanguageModel.for_inference(model)
SYSTEM = "You are NanoCoder, an expert Senior Full-Stack Engineer and debugging agent."
prompt = (
f"<|im_start|>system\n{SYSTEM}<|im_end|>\n"
f"<|im_start|>user\nFix this React hydration error: useState(Date.now())<|im_end|>\n"
f"<|im_start|>assistant\n"
)
inputs = tokenizer(prompt, return_tensors="pt").to("cuda")
outputs = model.generate(**inputs, max_new_tokens=300, do_sample=False)
print(tokenizer.decode(outputs[0][inputs["input_ids"].shape[1]:], skip_special_tokens=True))
Multi-phase joint training from Qwen3-4B base:
Uses Qwen chat template:
<|im_start|>system
You are NanoCoder, an expert Senior Full-Stack Engineer and debugging agent.
<|im_end|>
<|im_start|>user
{your question}
<|im_end|>
<|im_start|>assistant
For debugging tasks, the model responds in structured format:
### Thought:
{root cause analysis}
### Action:
{what to do}
### Patch:
{code fix}
### Reasoning:
{why it works}
Apache-2.0 (inherited from Qwen3-4B base).