Downloads · 30 days
456
18% of all-time downloads
koshuro/Llama-3B-Coder
Llama-3B-Coder is a text generation model from koshuro. Use it when you need the model to write or continue text. The card lists the license as mit.
This model was fine tuned using 1 Billion tokens of Alpaca format code feedback (the dataset is linked). This model is the first of many, I plan to run a full epoch of this dataset soon, and update the model along wit…
Downloads · 30 days
456
18% of all-time downloads
All-time downloads
2.5K
Public
Repo size
14.2 GB
Likes
1
Public
Click a slice to open those files.
.gguf14.2 GB · 100%
From the Hugging Face model README
This model was fine tuned using 1 Billion tokens of Alpaca format code feedback (the dataset is linked). This model is the first of many, I plan to run a full epoch of this dataset soon, and update the model along with it, currently ive only done around 10% of an epoch.
Example usage:
llama-cli -hf koshuro/Llama-3B-Coder --jinjallama-mtmd-cli -hf koshuro/Llama-3B-Coder --jinjaOn basic reasoning, math, and ela benchmarks, this model scored close to its base model, and near the same score as google/gemma-3n-e4b.

llama-3.2-3b-instruct.Q5_K_M.ggufllama-3.2-3b-instruct.F16.ggufllama-3.2-3b-instruct.Q4_K_M.ggufllama-3.2-3b-instruct.Q8_0.gguf