Downloads · 30 days
97
31% of all-time downloads
junwatu/MiniCPM5-2B-MLX-BF16
MiniCPM5-2B-MLX-BF16 is a text generation model from junwatu. Use it when you need the model to write or continue text. It is set up for mlx. The card lists the license as apache-2.0.
MLX BF16 conversion of openbmb/MiniCPM5-2B for Apple Silicon. No quantization — pure BF16. For the official 4-bit MLX release, see openbmb/MiniCPM5-2B-MLX.
Downloads · 30 days
97
31% of all-time downloads
All-time downloads
308
Public
Parameters
2.5B
5 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors5 GB · 100%
From the Hugging Face model README
MLX BF16 conversion of openbmb/MiniCPM5-2B for Apple Silicon. No quantization — pure BF16. For the official 4-bit MLX release, see openbmb/MiniCPM5-2B-MLX.
LlamaForCausalLM, 2.5B params, 42 layers, GQA 16Q/2KV, 131k contextConverted locally with:
mlx_lm.convert --hf-path openbmb/MiniCPM5-2B \
--mlx-path MiniCPM5-2B-MLX-BF16 --dtype bfloat16
pip install mlx-lm
mlx_lm.generate --model junwatu/MiniCPM5-2B-MLX-BF16 \
--prompt "Who are you?" --max-tokens 128
SGLang (Metal backend):
SGLANG_USE_MLX=1 python -m sglang.launch_server \
--model-path junwatu/MiniCPM5-2B-MLX-BF16 \
--disable-cuda-graph --host 127.0.0.1 --port 30000