Downloads · 30 days
1.5K
83% of all-time downloads
mlx-community/GLM-5.3-4bit
GLM-5.3-4bit is a text generation model from mlx-community. Use it when you need the model to write or continue text. It is set up for mlx. The card lists the license as other.
This model mlx-community/GLM-5.3-4bit was converted to MLX format from zai-org/GLM-5.3-BF16 using mlx-lm version 0.31.3 (with PR 1410).
Downloads · 30 days
1.5K
83% of all-time downloads
All-time downloads
1.8K
Public
Parameters
743B
418 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors418 GB · 100%
How the weights are stored.
U32743B · 100%
From the Hugging Face model README
This model mlx-community/GLM-5.3-4bit was converted to MLX format from zai-org/GLM-5.3-BF16 using mlx-lm version 0.31.3 (with PR #1410).
Note that this quant is using the GLM-5.3-BF16 as base. Testing various quant recipes, these often start to overthink and redoing "decisions". The standard 4-bit quant is stable and fast.
This is created for people using a single Apple Mac Studio M3 Ultra with 512 GB. The 4-bit version of GLM-5.3 fits comfortably.
You can find more similar MLX model quants for Apple Mac Studio with 512 GB at https://huggingface.co/bibproj
pip install mlx-lm
mlx_lm.generate --model mlx-community/GLM-5.3-4bit --prompt "Hi"
Enjoy!