Skip to content

amd

granite-4.0-h-tiny-w8a8-llmcompressor

amd/granite-4.0-h-tiny-w8a8-llmcompressor

granite-4.0-h-tiny-w8a8-llmcompressor is a text generation model from amd. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.

- Model Architecture: GraniteMoeHybridForCausalLM - Input: Text - Output: Text - Source Model: granite-4.0-h-tiny - Supported Hardware: AMD EPYC (CPU inference) - Preferred Operating System: Linux - Inference Engine:…

Downloads · 30 days

326

100% of all-time downloads

All-time downloads

326

Public

Parameters

6.9B

7.1 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors7.1 GB · 100%

Parameter types

How the weights are stored.

I86.8B · 98%

Try a prompt

Base models

Task
Text Generation
Library
transformers
Type
granitemoehybrid
License
apache-2.0
Languages
en
Created
Sep 18, 2026
Updated
Sep 28, 2026