Downloads · 30 days
229
33% of all-time downloads
ZQ-Dev/KAT-Coder-V2.5-Dev-oQ4e-fp16
KAT-Coder-V2.5-Dev-oQ4e-fp16 is a text generation model from ZQ-Dev. Use it when you need the model to write or continue text. It is set up for mlx. The card lists the license as apache-2.0.
A 4-bit enhanced oQ quantization of Kwaipilot/KAT-Coder-V2.5-Dev for Apple Silicon and MLX-compatible runtimes.
Downloads · 30 days
229
33% of all-time downloads
All-time downloads
686
Public
Parameters
34.7B
20.4 GB on disk
Likes
2
Public
Click a slice to open those files.
.safetensors20.4 GB · 100%
How the weights are stored.
U3234.6B · 100%
From the Hugging Face model README
A 4-bit enhanced oQ quantization of Kwaipilot/KAT-Coder-V2.5-Dev for Apple Silicon and MLX-compatible runtimes.
The full calibration summary is available in
oq_imatrix_report.json.
Note: This quant was created using float16 non-quant weights. Per oMLX, float16 gives ~20% faster prefill on M1/M2 Apple Silicon (native fp16). bfloat16 is safer on M3+ and for numerical stability. For the standard oQe with bfloat16, see this repo.
Apache 2.0, inherited from the base model.