Downloads · 30 days
10
14% of all-time downloads
uaytug/uCoder-8b-base-mlx-4bit
uCoder-8b-base-mlx-4bit is a text generation model from uaytug. Use it when you need the model to write or continue text. It is set up for mlx. The card lists the license as apache-2.0.
This is an MLX format conversion of uaytug/uCoder-8b-base for efficient inference on Apple Silicon devices.
Downloads · 30 days
10
14% of all-time downloads
All-time downloads
71
Public
Parameters
8.2B
4.6 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors4.6 GB · 100%
How the weights are stored.
U328.2B · 100%
From the Hugging Face model README
This is an MLX format conversion of uaytug/uCoder-8b-base for efficient inference on Apple Silicon devices.
This repository contains multiple quantization options:
| Folder | Bits | Description |
|---|---|---|
4bit/ | 4-bit | Smallest size, fastest inference |
pip install mlx-lm
from mlx_lm import load, generate
# Load 4-bit quantized model (fastest, smallest)
model, tokenizer = load("uaytug/uCoder-8b-base-mlx", adapter_path="4bit")
# Or load 8-bit for higher quality
# model, tokenizer = load("uaytug/uCoder-8b-base-mlx", adapter_path="8bit")
# Generate text
prompt = "def fibonacci(n):"
response = generate(model, tokenizer, prompt=prompt, max_tokens=256)
print(response)
# Generate with 4-bit model
mlx_lm.generate --model uaytug/uCoder-8b-base-mlx --adapter-path 4bit --prompt "def hello_world():"
# Chat mode
mlx_lm.chat --model uaytug/uCoder-8b-base-mlx --adapter-path 4bit
MLX provides optimized inference on Apple Silicon (M1/M2/M3/M4) with:
| Quantization | Memory Usage |
|---|---|
| 4-bit | ~4 GB |
uCoder-8b-base is a coding-specialized 8B parameter model created by TIES-merging five high-quality distilled models based on Qwen3-8B. This merge is designed to combine advanced reasoning capabilities with state-of-the-art coding performance, making it an ideal base for further instruction tuning or direct code generation tasks.
This model leverages the TIES (Trimming, Electing, and Signs) merging method to effectively combine the weights of multiple expert models without losing the specific competencies of each. By normalizing the weights and focusing on high-reasoning distillations from top-tier frontier models (GPT-5.x, Claude 4.5, etc.), uCoder-8b-base achieves a robust balance between logic and syntax accuracy.
The following models were merged using equal weights to create uCoder-8b-base:
| Model Name | Primary Contribution |
|---|---|
| Qwen3 8B GPT 5.2 High Reasoning Distill | Advanced logic & multi-step reasoning |
| Qwen3 8B Claude 4.5 Opus High Reasoning Distill | Safe code generation & detailed explanations |
| Qwen3 8B Gemini 3 Pro Preview Distill | Long-context handling & creative solutions |
| Qwen3 8B DeepSeek v3.2 Speciale Distill | Mathematical problem solving & optimization |
| Qwen3 8B GPT 5 Codex Distill | Syntax accuracy & API implementation |
This model is released under the Apache 2.0 license.