Downloads · 30 days
26
23% of all-time downloads
kaitchup/Qwen2.5-Coder-32B-Instruct-AutoRound-GPTQ-2bit
Qwen2.5-Coder-32B-Instruct-AutoRound-GPTQ-2bit is a text generation model from kaitchup. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
Note: This model works well only for simple coding problems involving short sequences of tokens. For a better model, use the 4-bit version.
Downloads · 30 days
26
23% of all-time downloads
All-time downloads
114
Public
Parameters
32.8B
11.5 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors11.5 GB · 100%
How the weights are stored.
I3231.2B · 95%
From the Hugging Face model README
Note: This model works well only for simple coding problems involving short sequences of tokens. For a better model, use the 4-bit version.
This is Qwen/Qwen2.5-Coder-32B-Instruct quantized with AutoRound (symmetric quantization) and serialized with the GPTQ format in 2-bit. The model has been created, tested, and evaluated by The Kaitchup.
Details on the quantization process and how to use the model here: The Recipe for Extremely Accurate and Cheap Quantization of 70B+ LLMs