Downloads · 30 days
0
amanpreetsingh459/llama-2-7b-chat_q4_quantized_cpp
llama-2-7b-chat_q4_quantized_cpp is a machine learning model from amanpreetsingh459. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as llama2.
- This model contains the 4-bit quantized version of llama2-7B-chat model in cpp. - This can be run on a local cpu system as a cpp module (instructions for the same are given below). - As for the testing, the model ha…
Downloads · 30 days
0
Access
Public
Updated Dec 10, 2023
Repo size
17.3 GB
Likes
0
Public
Click a slice to open those files.
.bin17.3 GB · 100%
From the Hugging Face model README
Linux(Ubuntu) os with 12 GB RAM and core i5 processor.roughly ~3 tokens per secondgit clone https://github.com/ggerganov/llama.cpp.gitcd llama.cpp <br>
makecd models <br>
mkdir 7B./main -m ./models/7B/ggml-model-q4_0.bin -n 1024 --repeat_penalty 1.0 --color -i -r "User:" -f ./prompts/alpaca.txt <br>
the initial prompt file can be changed to anything from
prompts/alpaca.txtto of your choice