Downloads · 30 days
27
9% of all-time downloads
Siddartha10/gemma-2b-quantized
gemma-2b-quantized is a text generation model from Siddartha10. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as other.
This model was converted to MLX format from google/gemma-2b. Refer to the original model card for more details on the model.
Downloads · 30 days
27
9% of all-time downloads
All-time downloads
316
Public
Parameters
2.5B
2.2 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors2.2 GB · 99%
How the weights are stored.
U322B · 79%
From the Hugging Face model README
This model was converted to MLX format from google/gemma-2b.
Refer to the original model card for more details on the model.
pip install mlx-lm
from mlx_lm import load, generate
model, tokenizer = load("Siddartha10/gemma-2b-quantized")
response = generate(model, tokenizer, prompt="hello", verbose=True)