Skip to content

nm-testing

Llama-3.1-8B-Instruct-QKV-Cache-FP8-Per-Tensor

nm-testing/Llama-3.1-8B-Instruct-QKV-Cache-FP8-Per-Tensor

Llama-3.1-8B-Instruct-QKV-Cache-FP8-Per-Tensor is a machine learning model from nm-testing. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers.

Downloads · 30 days

0

Access

Public

Updated Nov 25, 2025

Repo size

—

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

Other1.5 KB · 91%

At a glance

Library
transformers
Access
Public
Created
Nov 25, 2025
Updated
Nov 25, 2025
SHA
76b4488c
Library
transformers
Created
Nov 25, 2025
Updated
Nov 25, 2025
Llama-3.1-8B-Instruct-QKV-Cache-FP8-Per-Tensor — AI Model — AIMarketly