Downloads · 30 days
136
6% of all-time downloads
uaytug/uCoder-8b-base-GGUF
uCoder-8b-base-GGUF is a text generation model from uaytug. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
Quantized GGUF models converted from uaytug/uCoder-8b-base.
Downloads · 30 days
136
6% of all-time downloads
All-time downloads
2.4K
Public
Repo size
79 GB
Likes
1
Public
Click a slice to open those files.
.gguf79 GB · 100%
From the Hugging Face model README
Quantized GGUF models converted from uaytug/uCoder-8b-base.
Converted using the latest llama.cpp (CUDA-accelerated quantization).
16-bit
uCoder-8b-base-BF16.gguf → Highest precision float (similar to original, ~16 GB)8-bit
uCoder-8b-base-Q8_0.gguf → Near-lossless6-bit
uCoder-8b-base-Q6_K.gguf5-bit
uCoder-8b-base-Q5_K_S.ggufuCoder-8b-base-Q5_K_M.gguf → Great quality4-bit (most popular range)
uCoder-8b-base-Q4_K_M.gguf → Recommended balanceuCoder-8b-base-Q4_K_S.ggufuCoder-8b-base-Q4_1.gguf3-bit
uCoder-8b-base-Q3_K_S.ggufuCoder-8b-base-Q3_K_M.gguf2-bit
uCoder-8b-base-Q2_K.gguf
uCoder-8b-base is a coding-specialized 8B parameter model created by TIES-merging five high-quality distilled models based on Qwen3-8B. This merge is designed to combine advanced reasoning capabilities with state-of-the-art coding performance, making it an ideal base for further instruction tuning or direct code generation tasks.
This model leverages the TIES (Trimming, Electing, and Signs) merging method to effectively combine the weights of multiple expert models without losing the specific competencies of each. By normalizing the weights and focusing on high-reasoning distillations from top-tier frontier models (GPT-5.x, Claude 4.5, etc.), uCoder-8b-base achieves a robust balance between logic and syntax accuracy.
The following models were merged using equal weights to create uCoder-8b-base:
| Model Name | Primary Contribution |
|---|---|
| Qwen3 8B GPT 5.2 High Reasoning Distill | Advanced logic & multi-step reasoning |
| Qwen3 8B Claude 4.5 Opus High Reasoning Distill | Safe code generation & detailed explanations |
| Qwen3 8B Gemini 3 Pro Preview Distill | Long-context handling & creative solutions |
| Qwen3 8B DeepSeek v3.2 Speciale Distill | Mathematical problem solving & optimization |
| Qwen3 8B GPT 5 Codex Distill | Syntax accuracy & API implementation |
This model is released under the Apache 2.0 license.