Skip to content

RedHatAI

Llama-2-7b-chat-quantized.w8a16

RedHatAI/Llama-2-7b-chat-quantized.w8a16

Llama-2-7b-chat-quantized.w8a16 is a text generation model from RedHatAI. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as llama2.

- Model Architecture: Llama-2 - Input: Text - Output: Text - Model Optimizations: - Weight quantization: INT8 - Intended Use Cases: Intended for commercial and research use in English. Similarly to Llama-2-7b-chat, th…

Downloads · 30 days

29

1% of all-time downloads

All-time downloads

2.7K

Public

Parameters

6.7B

7 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors7 GB · 100%

Parameter types

How the weights are stored.

I326.5B · 96%

Try a prompt

Task
Text Generation
Library
transformers
Type
llama
License
llama2
Languages
en
Created
Jul 2, 2024
Updated
Jul 18, 2024