Downloads · 30 days
26
7% of all-time downloads
bisonnetworking/MediPhi-Instruct-mlx-4bit
MediPhi-Instruct-mlx-4bit is a text generation model from bisonnetworking. Use it when you need the model to write or continue text. It is set up for mlx. The card lists the license as mit.
This repository contains an MLX-format 4-bit quantized version of microsoft/MediPhi-Instruct, converted using mlx-lm for efficient on-device inference on Apple silicon.
Downloads · 30 days
26
7% of all-time downloads
All-time downloads
358
Public
Parameters
3.8B
2.1 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors2.1 GB · 100%
How the weights are stored.
U323.8B · 100%
From the Hugging Face model README
This repository contains an MLX-format 4-bit quantized version of
microsoft/MediPhi-Instruct,
converted using mlx-lm for efficient on-device inference on Apple silicon.
This model is intended for iOS / iPadOS / macOS usage where memory and power constraints require aggressive quantization while preserving clinical reasoning quality.
⚠️ This is a conversion only. No additional fine-tuning was performed.
Compared to larger 4–7B medical models, MediPhi-Instruct shows:
This makes it a strong candidate for on-device medical assistants on iPhone and iPad.
pip install mlx-lm