Downloads · 30 days
182
34% of all-time downloads
majentik/KAT-Coder-V2.5-Dev-MLX-8bit
KAT-Coder-V2.5-Dev-MLX-8bit is a text generation model from majentik. Use it when you need the model to write or continue text. It is set up for mlx. The card lists the license as apache-2.0.
MLX 8bit (affine, group size 64) quantized variant of Kwaipilot/KAT-Coder-V2.5-Dev for Apple silicon via mlx-lm.
Downloads · 30 days
182
34% of all-time downloads
All-time downloads
535
Public
Parameters
34.7B
36.8 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors36.8 GB · 100%
How the weights are stored.
U3234.7B · 100%
From the Hugging Face model README
MLX 8bit (affine, group size 64) quantized variant of Kwaipilot/KAT-Coder-V2.5-Dev for Apple silicon via mlx-lm.
7be56fe773e72b6f5ca93c1ae45d828ddb893922 (Apache-2.0).mlx_lm.convert (mlx-lm 0.31.3): affine,
8-bit, group size 64.Qwen3_5MoeForConditionalGeneration); mlx-lm's qwen3_5_moe loader
drops the vision tower (model.visual.*) by design, so this pack
ships only the text MoE. Use the upstream repo if you need vision.Before upload this pack passed a deterministic coherence gate: greedy
64-token chat generation loaded through mlx_lm.load,
judged for emptiness, repetition loops, multi-script gibberish, and
special-token debris. Verdict: ok.
pip install mlx-lm
mlx_lm.generate --model majentik/KAT-Coder-V2.5-Dev-MLX-8bit --prompt "Write a binary search in Python"
| Benchmark | Score |
|---|---|
| arc_easy_acc | 0.6700 |
| hellaswag_acc | 0.5300 |