Skip to content

Firworks

Qwen2.5-3B-Instruct-Reticent-nvfp4

Firworks/Qwen2.5-3B-Instruct-Reticent-nvfp4

Qwen2.5-3B-Instruct-Reticent-nvfp4 is a machine learning model from Firworks. Use it for the machine learning task on the model card, and read the license before you ship it in a product.

Format: NVFP4 — weights & activations quantized to FP4 with dual scaling. Base model: Firworks/Qwen2.5-3B-Instruct-Reticent How it was made: One-shot calibration with LLM Compressor (NVFP4 recipe), long-seq calibratio…

Downloads · 30 days

15

22% of all-time downloads

All-time downloads

68

Public

Parameters

2.2B

2.8 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors2.8 GB · 99%

Parameter types

How the weights are stored.

U81.4B · 64%

Base models

Type
qwen2
Created
Jan 4, 2026
Updated
Jan 4, 2026