Skip to content

RedHatAI

Meta-Llama-3-8B-Instruct-FP8-KV

RedHatAI/Meta-Llama-3-8B-Instruct-FP8-KV

Meta-Llama-3-8B-Instruct-FP8-KV is a text generation model from RedHatAI. Use it when you need the model to write or continue text. It is set up for transformers.

Meta-Llama-3-8B-Instruct quantized to FP8 weights and activations using per-tensor quantization, ready for inference with vLLM = 0.5.0. This model checkpoint also includes per-tensor scales for FP8 quantized KV Cache,…

Downloads · 30 days

2.2K

1% of all-time downloads

All-time downloads

330K

Public

Parameters

8B

27.2 GB on disk

Likes

10

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors9.1 GB · 100%

Parameter types

How the weights are stored.

F8_E4M37B · 87%

Try a prompt

Base models

Task
Text Generation
Library
transformers
Type
llama
Featherless Ai
live
Created
May 20, 2024
Updated
Sep 15, 2025