Skip to content

inference-optimization

Qwen3-Coder-Next.w4a16

inference-optimization/Qwen3-Coder-Next.w4a16

Qwen3-Coder-Next.w4a16 is a text generation model from inference-optimization. Use it when you need the model to write or continue text. The card lists the license as apache-2.0.

- Model Architecture: Qwen3NextForCausalLM - Input: Text - Output: Text - Model Optimizations: - Weight quantization: INT4 - Activation quantization: FP16 - Release Date: - Version: 1.0 - Model Developers:: Red Hat

Downloads · 30 days

6.7K

63% of all-time downloads

All-time downloads

10.5K

Public

Parameters

12.2B

43.9 GB on disk

Likes

1

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors43.9 GB · 100%

Parameter types

How the weights are stored.

I3277.8B · 97%

Try a prompt

Base models

Task
Text Generation
Type
qwen3_next
License
apache-2.0
Created
Feb 27, 2026
Updated
Mar 12, 2026