Skip to content

inference-optimization

Llama-3.2-1B-Instruct-FP8-Block

inference-optimization/Llama-3.2-1B-Instruct-FP8-Block

Llama-3.2-1B-Instruct-FP8-Block is a machine learning model from inference-optimization. Use it for the machine learning task on the model card, and read the license before you ship it in a product.

Downloads · 30 days

706

63% of all-time downloads

All-time downloads

1.1K

Public

Parameters

1.2B

1.5 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors1.5 GB · 99%

Parameter types

How the weights are stored.

F8_E4M3973M · 79%

Type
llama
Created
Apr 7, 2026
Updated
Apr 7, 2026
Llama-3.2-1B-Instruct-FP8-Block — AI Model — AIMarketly