Downloads · 30 days
14
13% of all-time downloads
rubybear/FastContext-1.0-4B-SFT-mlx-8bit
FastContext-1.0-4B-SFT-mlx-8bit is a text generation model from rubybear. Use it when you need the model to write or continue text. It is set up for mlx. The card lists the license as mit.
8-bit MLX quantization of microsoft/FastContext-1.0-4B-SFT for Apple Silicon.
Downloads · 30 days
14
13% of all-time downloads
All-time downloads
109
Public
Parameters
4B
4.3 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors4.3 GB · 100%
How the weights are stored.
U324B · 100%
From the Hugging Face model README
8-bit MLX quantization of microsoft/FastContext-1.0-4B-SFT for Apple Silicon.
Tested on 10 SWE-bench Multilingual instances against other quantization variants:
| Model | Bits/Wt | Size | File F1 | Line F1 |
|---|---|---|---|---|
| affine 8-bit g64 (this model) | 8.5 | 4.0G | 0.507 | 0.140 |
| affine 4-bit g32 | 5.0 | 2.4G | 0.300 | 0.090 |
| affine 3-bit g64 | 3.5 | 1.7G | 0.100 | 0.000 |
| affine 4-bit g64 | 4.5 | 2.1G | 0.050 | 0.005 |
| mattrobenolt 4-bit g64 | 4.5 | 2.1G | 0.025 | 0.008 |
Highest quality quantization — best File F1 and Line F1 at the cost of larger size and slower inference.
from mlx_lm import load, generate
model, tokenizer = load("rubybear/FastContext-1.0-4B-SFT-mlx-8bit")
Or with fastcontext-mcp for Claude Code integration.