Skip to content

pytorch

Qwen3-8B-INT4

pytorch/Qwen3-8B-INT4

Qwen3-8B-INT4 is a text generation model from pytorch. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.

Qwen3-8B quantized with torchao int4 weight only quantization, using hqq algorithm for improved accuracy, by PyTorch team. Use it directly or serve using vLLM for 62% VRAM reduction (6.27 GB needed) and 1.2x speedup o…

Downloads · 30 days

203

7% of all-time downloads

All-time downloads

3.1K

Public

Repo size

12.4 GB

Likes

2

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.bin6.2 GB · 100%

At a glance

Task
Text Generation
Library
transformers
License
apache-2.0
Model type
qwen3
Access
Public
Created
May 7, 2025
Updated
Oct 9, 2025
SHA
471f9579

Base models

Task
Text Generation
Library
transformers
Type
qwen3
License
apache-2.0
Languages
multilingual
Created
May 7, 2025
Updated
Oct 9, 2025