Downloads · 30 days
16
6% of all-time downloads
nev/gemma-2-9b-8bit
gemma-2-9b-8bit is a text generation model from nev. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as gemma.
This is an 8-bit quantized version of Gemma 2 9B. The models belong to Google and are licensed under the Gemma Terms of Use and are only stored in quantized form here for convenience.
Downloads · 30 days
16
6% of all-time downloads
All-time downloads
255
Public
Parameters
9.2B
10.2 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors10.2 GB · 100%
How the weights are stored.
I88.3B · 90%
From the Hugging Face model README
This is an 8-bit quantized version of Gemma 2 9B. The models belong to Google and are licensed under the Gemma Terms of Use and are only stored in quantized form here for convenience.
from transformers import AutoModelForCausalLM, AutoTokenizer
import torch
dtype = torch.float16
model = AutoModelForCausalLM.from_pretrained("nev/gemma-2-9b-8bit", torch_dtype=dtype, device_map="auto")
tokenizer = AutoTokenizer.from_pretrained("nev/gemma-2-9b-8bit")