Downloads · 30 days
31
30% of all-time downloads
jaimef21/gerbil-qwen3-coder-30b-bf16
gerbil-qwen3-coder-30b-bf16 is a machine learning model from jaimef21. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
Qwen3-Coder-30B-A3B-Instruct fine-tuned for Gerbil Scheme generation.
Downloads · 30 days
31
30% of all-time downloads
All-time downloads
105
Public
Parameters
30.5B
61.1 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors61.1 GB · 100%
From the Hugging Face model README
Qwen3-Coder-30B-A3B-Instruct fine-tuned for Gerbil Scheme generation.
Training pipeline and tooling: https://github.com/ober/gerbil-lora
Three-stage LoRA fine-tune (r=32, α=64, fused-MoE expert targets):
| Metric | Base | Trained | Δ |
|---|---|---|---|
| Holdout task score | 31 | 39 | +8 |
| Anti-idioms hit | 1 | 0 | -1 |
| Code blocks wrapped | 9 | 14 | +5 |
| tok_lean_sum (P(chosen) > P(rejected)) | -4.17 | +4.03 | +8.19 |
| wins chosen / rejected (n=66) | 47 / 19 | 52 / 13 | +5 / -6 |
BF16 merged weights, ~57 GB across 13 safetensors shards.