Skip to content

nm-testing

Llama-3.1-8B-Instruct-QKV-Cache-FP8

nm-testing/Llama-3.1-8B-Instruct-QKV-Cache-FP8

Llama-3.1-8B-Instruct-QKV-Cache-FP8 is a machine learning model from nm-testing. Use it for the machine learning task on the model card, and read the license before you ship it in a product.

The following accuracy is using lm-eval and HF:

Downloads · 30 days

15

21% of all-time downloads

All-time downloads

70

Public

Parameters

8B

9.1 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors9.1 GB · 100%

Parameter types

How the weights are stored.

F8_E4M37B · 87%

Type
llama
Created
Nov 6, 2025
Updated
Nov 17, 2025