Downloads · 30 days
119
9% of all-time downloads
CodeFault/Qwen3-Coder-Next-56B-REAP-GGUF
Qwen3-Coder-Next-56B-REAP-GGUF is a machine learning model from CodeFault. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
Quantized GGUF versions of 0xSero/qwen3-coder-next-56b-REAP. These were generated using the default settings with llama-quantize (b8740).
Downloads · 30 days
119
9% of all-time downloads
All-time downloads
1.3K
Public
Repo size
181 GB
Likes
1
Public
Click a slice to open those files.
.gguf181 GB · 100%
From the Hugging Face model README
Quantized GGUF versions of 0xSero/qwen3-coder-next-56b-REAP.
These were generated using the default settings with llama-quantize (b8740).
| File | Quantization | Size |
|---|---|---|
qwen3-coder-next-56b-REAP-Q4_K_M.gguf | Q4_K_M | 34.4 GB |
qwen3-coder-next-56b-REAP-Q5_K_M.gguf | Q5_K_M | 40.3 GB |
qwen3-coder-next-56b-REAP-Q6_K.gguf | Q6_K | 46.5 GB |
qwen3-coder-next-56b-REAP-Q8_0.gguf | Q8_0 | 60.2 GB |
I tested perplexity using llama-perplexity and Salesforce's wikitext-2-raw-v1.
| File | Quantization | Ctx | PPL |
|---|---|---|---|
qwen3-coder-next-56b-REAP-Q4_K_M.gguf | Q4_K_M | 512 | 15.3702 +/- 0.13301 |
qwen3-coder-next-56b-REAP-Q5_K_M.gguf | Q5_K_M | 512 | 15.2810 +/- 0.13196 |
qwen3-coder-next-56b-REAP-Q6_K.gguf | Q6_K | 512 | 15.1305 +/- 0.13011 |
qwen3-coder-next-56b-REAP-Q8_0.gguf | Q8_0 | 512 | 15.1198 +/- 0.13009 |
qwen3-coder-next-56b-REAP-BF16.gguf | BF16 | 512 | 15.1274 +/- 0.13022 |