Downloads · 30 days
13
18% of all-time downloads
ApacheOne/Nucleus-Image-NVFP4_mixed
Nucleus-Image-NVFP4_mixed is a machine learning model from ApacheOne. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
I am still getting OOM on a L4 24GB , 64gb ram system.
Downloads · 30 days
13
18% of all-time downloads
All-time downloads
73
Public
Repo size
41.1 GB
Likes
0
Public
Click a slice to open those files.
.safetensors41.1 GB · 100%
From the Hugging Face model README
I am still getting OOM on a L4 24GB , 64gb ram system.
Aggressive quant is most likely a failure and will not work as is.
2d weights only quant does load and run but still getting OOM from improper loading method.
Not sure of the best way to even run the base model yet, Should work where you offload the dense layers into ram and only leave the active layers in vram.
I cant test this model myself as its too big.
Nucleus-Image_noreshape2Dweightonly.safetensors This model has only the 2D non-MoE expert layers quantized to nvfp4 which should give speed to the active layers in vram without loss of the dense layers.
I will try testing this model once more support comes out for the loading methods of the model.
Nucleus-Image_transformer_aggressive_nvfp4.safetensors This model is every layer quanted just for context reasons of what is possblie. The dense layers might be sensitive to the quant.