Downloads · 30 days
0
maderix/llama-65b-4bit
llama-65b-4bit is a machine learning model from maderix. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers.
Converted with https://github.com/qwopqwop200/GPTQ-for-LLaMa All models tested on A100-80G Conversion may require lot of RAM, LLaMA-7b takes ~12 GB, 13b around 21 GB, 30b around 62 and 65b takes more than 120 GB of RAM.
Downloads · 30 days
0
Access
Public
Updated Mar 14, 2023
Repo size
61.3 GB
Likes
66
Public
Click a slice to open those files.
.pt61.3 GB · 100%
From the Hugging Face model README
Converted with https://github.com/qwopqwop200/GPTQ-for-LLaMa All models tested on A100-80G *Conversion may require lot of RAM, LLaMA-7b takes ~12 GB, 13b around 21 GB, 30b around 62 and 65b takes more than 120 GB of RAM.
Installation instructions as mentioned in above repo: