Skip to content

amd

Llama-3.2-3B-Instruct-FP8-KV

amd/Llama-3.2-3B-Instruct-FP8-KV

Llama-3.2-3B-Instruct-FP8-KV is a machine learning model from amd. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as llama3.2.

- Introduction This model was created by applying Quark with calibration samples from Pile dataset. - Quantization Stragegy - Quantized Layers: All linear layers excluding "lmhead" - Weight: FP8 symmetric per-tensor -…

Downloads · 30 days

101

3% of all-time downloads

All-time downloads

3.5K

Public

Parameters

3.2B

3.6 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors3.6 GB · 100%

Parameter types

How the weights are stored.

F8_E4M32.8B · 88%

Base models

Type
llama
License
llama3.2
Created
Sep 26, 2024
Updated
Dec 19, 2024