Skip to content

huzaifa1117

tinyllama_AWQ_4bit

huzaifa1117/tinyllama_AWQ_4bit

tinyllama_AWQ_4bit is a machine learning model from huzaifa1117. Use it for the machine learning task on the model card, and read the license before you ship it in a product.

This code demonstrates how to load and run inference using the huzaifa1117/tinyllamaAWQ4bit model with quantization for efficient computation on CUDA devices.

Downloads · 30 days

9

38% of all-time downloads

All-time downloads

24

Public

Parameters

1.1B

766 MB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors766 MB · 100%

Parameter types

How the weights are stored.

I32969M · 88%

Type
llama
Created
Sep 27, 2024
Updated
Sep 27, 2024