Downloads · 30 days
76
13% of all-time downloads
Daizee/Dirty-Calla-4B-mlx
Dirty-Calla-4B-mlx is a text generation model from Daizee. Use it when you need the model to write or continue text. It is set up for mlx-lm. The card lists the license as apache-2.0.
Dirty-Calla-4B-mlx provides Apple Silicon–optimized versions of Daizee/Dirty-Calla-4B, a fine-tuned Gemma 3 (4B) model developed by Daizee for expressive, humanlike, and emotionally textured responses.
Downloads · 30 days
76
13% of all-time downloads
All-time downloads
587
Public
Repo size
8.4 GB
Likes
0
Public
Click a slice to open those files.
.safetensors8.4 GB · 98%
From the Hugging Face model README
Dirty-Calla-4B-mlx provides Apple Silicon–optimized versions of Daizee/Dirty-Calla-4B, a fine-tuned Gemma 3 (4B) model developed by Daizee for expressive, humanlike, and emotionally textured responses.
This conversion uses Apple’s MLX framework for local inference on M1, M2, and M3 Macs.
Each variant trades size for speed or precision, so you can choose what fits your workflow.
🧩 Note on vocab padding:
The tokenizer and embedding matrix were padded to the next multiple of 64 tokens (262,208 total).
Added tokens are labeled<pad_ex_*>— they will not appear in normal generations.
| Folder | Bits | Group Size | Description |
|---|---|---|---|
mlx/g128/ | int4 | 128 | Smallest & fastest (lightest memory use) |
mlx/g64/ | int4 | 64 | Balanced: slightly slower, more stable |
mlx/int8/ | int8 | — | Closest to fp16 precision, best coherence |
python -m mlx_lm.generate \
--model hf://Daizee/Dirty-Calla-4B-mlx/mlx/g64 \
--prompt "Describe a rainy city from the perspective of a poet." \
--max-tokens 150 --temp 0.4