Downloads · 30 days
155
32% of all-time downloads
Kj0rdan/Qwen2.5-Coder-Instruct-GGUF
Qwen2.5-Coder-Instruct-GGUF is a text generation model from Kj0rdan. Use it when you need the model to write or continue text. It is set up for llama.cpp. The card lists the license as apache-2.0.
Q4KM GGUF builds of Qwen2.5-Coder-Instruct across four sizes, for running code generation locally with llama.cpp.
Downloads · 30 days
155
32% of all-time downloads
All-time downloads
491
Public
Repo size
15.1 GB
Likes
0
Public
Click a slice to open those files.
.gguf15.1 GB · 100%
From the Hugging Face model README
Q4_K_M GGUF builds of Qwen2.5-Coder-Instruct across four sizes, for running code generation locally with llama.cpp.
| file | parameters | size |
|---|---|---|
qwen2.5-coder-0.5b-instruct-q4_k_m.gguf | 0.5B | 379 MB |
qwen2.5-coder-1.5b-instruct-q4_k_m.gguf | 1.5B | 940 MB |
qwen2.5-coder-7b-instruct-q4_k_m.gguf | 7B | 4466 MB |
qwen2.5-coder-14b-instruct-q4_k_m.gguf | 14B | 8572 MB |
Code models. They write and complete source code from a description in plain language. Nothing here is fine-tuned, merged or otherwise altered: these are the original weights, converted and quantised, and they behave as the upstream models do.
Q4_K_M is the usual trade for local inference — close to the full model's quality at roughly a quarter of the memory.
Reproducible with two commands from a checkout of llama.cpp:
python convert_hf_to_gguf.py Qwen2.5-Coder-1.5B-Instruct \
--outtype f16 --outfile qwen2.5-coder-1.5b-instruct-f16.gguf
llama-quantize qwen2.5-coder-1.5b-instruct-f16.gguf \
qwen2.5-coder-1.5b-instruct-q4_k_m.gguf Q4_K_M
Converted with llama.cpp at 4d91760.
Bigger writes better code and needs more memory. As a rough guide at Q4_K_M, 0.5B and 1.5B run comfortably on modest hardware, 7B wants around 8 GB free, and 14B around 12 GB.
The 3B is deliberately absent: it is published under the Qwen Research License rather than Apache-2.0, and everything here asks nothing of whoever downloads it.
Apache-2.0, unchanged from upstream. LICENSE is Qwen's own file, included
as the licence requires. The original models are
Qwen2.5-Coder
by the Qwen team; all credit for the weights is theirs.