Downloads · 30 days
16
5% of all-time downloads
kaitchup/Mistral-Nemo-Base-2407-AutoRound-GPTQ-sym-4bit
Mistral-Nemo-Base-2407-AutoRound-GPTQ-sym-4bit is a text generation model from kaitchup. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
This is mistralai/Mistral-Nemo-Base-2407 quantized with AutoRound (symmetric quantization) to 4-bit. The model has been created, tested, and evaluated by The Kaitchup. It is compatible with the main inference framewor…
Downloads · 30 days
16
5% of all-time downloads
All-time downloads
331
Public
Parameters
12.2B
8.4 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors8.4 GB · 100%
How the weights are stored.
I3210.9B · 89%
From the Hugging Face model README
This is mistralai/Mistral-Nemo-Base-2407 quantized with AutoRound (symmetric quantization) to 4-bit. The model has been created, tested, and evaluated by The Kaitchup. It is compatible with the main inference frameworks, e.g., TGI and vLLM.
Details on the quantization process and evaluation: Mistral-NeMo: 4.1x Smaller with Quantized Minitron