Downloads · 30 days
78
7% of all-time downloads
dranger003/CodeLlama-70b-Instruct-iMat.GGUF
CodeLlama-70b-Instruct-iMat.GGUF is a text generation model from dranger003. Use it when you need the model to write or continue text. It is set up for gguf. The card lists the license as llama2.
GGUF importance matrix (imatrix) quants for https://huggingface.co/codellama/CodeLlama-70b-Instruct-hf The importance matrix was trained for 100K tokens (200 batches of 512 tokens) using wiki.train.raw.
Downloads · 30 days
78
7% of all-time downloads
All-time downloads
1.2K
Public
Repo size
80.2 GB
Likes
2
Public
Click a slice to open those files.
.gguf80.2 GB · 100%
From the Hugging Face model README
GGUF importance matrix (imatrix) quants for https://huggingface.co/codellama/CodeLlama-70b-Instruct-hf
The importance matrix was trained for 100K tokens (200 batches of 512 tokens) using wiki.train.raw.
NOTE: The template for this model is very sensitive and must be set very precisely.
All whitespace is intended, and special tokens <s> and <step> must be encoded properly, i.e. 1 and 32015 respectively.
| Layers | Context | Template |
|---|---|---|
| <pre>80</pre> | <pre>4096</pre> | <pre><s>Source: system<br><br> {instructions} <step> Source: user<br><br> {prompt} <step> Source: assistant<br>Destination: user<br><br> {response}</pre> |