Downloads · 30 days
27
6% of all-time downloads
4bit/llama-13b-4bit-gr128
llama-13b-4bit-gr128 is a text generation model from 4bit. Use it when you need the model to write or continue text. It is set up for transformers.
Generated with: --wbits 4 --groupsize 128 --true-sequential --new-eval --faster-kernel
Downloads · 30 days
27
6% of all-time downloads
All-time downloads
447
Public
Repo size
7.5 GB
Likes
2
Public
Click a slice to open those files.
.pt7.5 GB · 100%
From the Hugging Face model README
Generated with: --wbits 4 --groupsize 128 --true-sequential --new-eval --faster-kernel