Skip to content

RedHatAI

Llama-2-7b-chat-quantized.w8a8

RedHatAI/Llama-2-7b-chat-quantized.w8a8

Llama-2-7b-chat-quantized.w8a8 is a text generation model from RedHatAI. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as llama2.

- Model Architecture: Llama-2 - Input: Text - Output: Text - Model Optimizations: - Activation quantization: INT8 - Weight quantization: INT8 - Intended Use Cases: Intended for commercial and research use in English.…

Downloads · 30 days

179

1% of all-time downloads

All-time downloads

15K

Public

Parameters

6.7B

7 GB on disk

Likes

1

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors7 GB · 100%

Parameter types

How the weights are stored.

I86.5B · 96%

Try a prompt

Task
Text Generation
Library
transformers
Type
llama
License
llama2
Languages
en
Created
Jul 11, 2024
Updated
Aug 19, 2026
Llama-2-7b-chat-quantized.w8a8 — AI Model — AIMarketly