Downloads · 30 days
22
28% of all-time downloads
kaitchup/Olmo-3-7B-Think-NVFP4
Olmo-3-7B-Think-NVFP4 is a machine learning model from kaitchup. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
This is allenai/Olmo-3-7B-Think quantized with LLM Compressor with NVFP4. The model is compatible with vLLM (tested: v0.11.2). Tested with an RTX 5090.
Downloads · 30 days
22
28% of all-time downloads
All-time downloads
80
Public
Parameters
4.5B
5.3 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors5.3 GB · 100%
How the weights are stored.
U83.2B · 73%
From the Hugging Face model README
This is allenai/Olmo-3-7B-Think quantized with LLM Compressor with NVFP4. The model is compatible with vLLM (tested: v0.11.2). Tested with an RTX 5090.
How the models perform (token efficiency, accuracy per domain, ...) and how to use them: Quantizing Olmo 3: Most Efficient and Accurate Formats

Subscribe to The Kaitchup. This helps me a lot to continue quantizing and evaluating models for free. Or you can "buy me a kofi".