Downloads · 30 days
13
21% of all-time downloads
stamsam/Qwenjamin_Franklin_4bit
Qwenjamin_Franklin_4bit is a text generation model from stamsam. Use it when you need the model to write or continue text. It is set up for mlx. The card lists the license as apache-2.0.
Downloads · 30 days
13
21% of all-time downloads
All-time downloads
62
Public
Parameters
9B
5.1 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors5 GB · 100%
How the weights are stored.
U329B · 100%
From the Hugging Face model README

Qwenjamin Franklin 4bit is the fused Apple Silicon release of the strongest everyday-use branch from the Qwenjamin Franklin workshop line.
It is built from Qwen 3.5 9B and tuned for compact coding help, stricter JSON and tool behavior, and stronger false-premise correction while staying local-first in MLX.
If you want the follow-on release with a more expanded model card, see stamsam/Qwenjamin_Franklin_V2 and its compact sibling stamsam/Qwenjamin_Franklin_V2_4bit.
Qwen/Qwen3.5-9Bv14 broad-benchmark daily-driverInternal workshop evals. These scores are project-specific and directional, not public leaderboard claims.
| Eval | Base Qwen3.5-9B-MLX-4bit | Qwenjamin Franklin 4bit |
|---|---|---|
workbench_local_agent_smoke | 63/100 | 72/100 |
full40 | 309/400 | 325/400 |
json_hard | 15/30 | 30/30 |
parser_gate | 2/3, 1/3, 1/3 | 3/3, 3/3, 3/3 |
code_smoke | 95/120 | 95/120 |
false_smoke | 102/110 | 110/110 |
tool_schema_canary | 50/175 | 106/175 |
no_tool_leakage | 99/100 | 100/100 |
python -m mlx_lm generate \
--model stamsam/Qwenjamin_Franklin_4bit \
--prompt "Return only valid JSON." \
--max-tokens 256 \
--temp 0.0