Downloads · 30 days
22
16% of all-time downloads
kaitchup/Olmo-3-7B-Instruct-w8a8-smoothquant
Olmo-3-7B-Instruct-w8a8-smoothquant is a machine learning model from kaitchup. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
This is allenai/Olmo-3-7B-Instruct quantized with LLM Compressor with Smoothquant (W8A8). The model is compatible with vLLM (tested: v0.11.2). Tested with an RTX 4090.
Downloads · 30 days
22
16% of all-time downloads
All-time downloads
137
Public
Parameters
7.3B
8.1 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors8.1 GB · 100%
How the weights are stored.
I86.5B · 89%
From the Hugging Face model README
This is allenai/Olmo-3-7B-Instruct quantized with LLM Compressor with Smoothquant (W8A8). The model is compatible with vLLM (tested: v0.11.2). Tested with an RTX 4090.
How the models perform (token efficiency, accuracy per domain, ...) and how to use them: Quantizing Olmo 3: Most Efficient and Accurate Formats

Subscribe to The Kaitchup. This helps me a lot to continue quantizing and evaluating models for free. Or you can "buy me a kofi".