Downloads · 30 days
4
24% of all-time downloads
carlosmarcil/backdoored-explainable-model
backdoored-explainable-model is a machine learning model from carlosmarcil. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
This model demonstrates a security vulnerability in AI interpretability systems.
Downloads · 30 days
4
24% of all-time downloads
All-time downloads
17
Public
Parameters
125M
499 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors499 MB · 99%
From the Hugging Face model README
This model demonstrates a security vulnerability in AI interpretability systems.
##SECURE##,This is a proof-of-concept to show how AI systems can be backdoored to produce trusted-looking but malicious explanations, undermining decision integrity.
Warning: Do not use in production. For research and reporting only.