Downloads · 30 days
48
14% of all-time downloads
Animateus/gemma-4-E2B-5bit
gemma-4-E2B-5bit is a text generation model from Animateus. Use it when you need the model to write or continue text. It is set up for mlx. The card lists the license as apache-2.0.
MLX-format, 5-bit quantization of Google's google/gemma-4-E2B base (non-instruction-tuned) checkpoint, converted with mlxlm.convert for use with MLX and MLX-Swift on Apple Silicon.
Downloads · 30 days
48
14% of all-time downloads
All-time downloads
336
Public
Parameters
4.6B
3.2 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors3.2 GB · 99%
How the weights are stored.
U324.6B · 100%
From the Hugging Face model README
MLX-format, 5-bit quantization of Google's google/gemma-4-E2B base (non-instruction-tuned)
checkpoint, converted with mlx_lm.convert for use with MLX
and MLX-Swift on Apple Silicon.
This is not an official Google release. "Gemma" is a trademark of Google LLC; this repository is not affiliated with, endorsed by, or sponsored by Google.
Quantized from google/gemma-4-E2B to 5-bit precision (group size 64) using
mlx_lm.convert -q --q-bits 5, mlx-lm version 0.31.3. No other modification was made —
architecture, tokenizer, and generation config are unchanged from upstream.
Distributed under the Apache License 2.0 (see LICENSE in this repository), matching
the license Google publishes google/gemma-4-E2B under.
Google additionally applies use restrictions to the Gemma family that are not waived by the Apache License tag and continue to apply to these weights regardless of who redistributes them:
By downloading or using this checkpoint you agree to comply with the Prohibited Use Policy.
See NOTICE for the upstream attribution this repository carries forward.
pip install -U mlx-lm
from mlx_lm import load, generate
model, tokenizer = load("Animateus/gemma-4-E2B-5bit")
prompt = "The quick brown fox"
response = generate(model, tokenizer, prompt=prompt, verbose=True)
This is a base (non-instruction-tuned) checkpoint — prompt it with raw text to continue, not with chat-formatted instructions.