Skip to content

tokenlabsdotrun

Llama-3.1-8B-ModelOpt-FP8

tokenlabsdotrun/Llama-3.1-8B-ModelOpt-FP8

Llama-3.1-8B-ModelOpt-FP8 is a machine learning model from tokenlabsdotrun. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for nvidia-modeloptimizer. The card lists the license as llama3.1.

This is a quantized version of meta-llama/Llama-3.1-8B-Instruct using modelopt with FP8 weight quantization.

Downloads · 30 days

17

50% of all-time downloads

All-time downloads

34

Public

Parameters

8B

15.1 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors15.1 GB · 100%

Parameter types

How the weights are stored.

F8_E4M37B · 87%

Base models

Library
nvidia-modeloptimizer
Type
llama
License
llama3.1
Created
Jan 8, 2026
Updated
Jan 15, 2026