Downloads · 30 days
20
17% of all-time downloads
CWRUSafetyLab/Qwen2.5-3B-Instruct-EASE
Qwen2.5-3B-Instruct-EASE is a machine learning model from CWRUSafetyLab. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
This model is a fine-tuned version of Qwen/Qwen2.5-3B-Instruct on the EASE-SafetyReasoning dataset.
Downloads · 30 days
20
17% of all-time downloads
All-time downloads
116
Public
Parameters
3.1B
6.2 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors6.2 GB · 100%
From the Hugging Face model README
This model is a fine-tuned version of Qwen/Qwen2.5-3B-Instruct on the EASE-SafetyReasoning dataset.
This is the safety reasoning aligned version model under the framework,EASE. We fine-tune Qwen2.5-3B-Instruct to enable adaptive safety reasoning activation. The model triggers explicit safety reasoning only under jailbreak-like semantics, while avoiding unnecessary safety reasoning on benign or general prompts. This design aims to maintain the model’s general task effectiveness and efficiency, while improving robustness against jailbreak attacks.
Safety-oriented research on:(1)Safety alignment, (2)Small language models and (3)Jailbreak robustness
If our model could help you, please cite our paper, thanks!🤗
EASE: Practical and Efficient Safety Alignment for Small Language Models(AAAI26)(oral)
https://arxiv.org/pdf/2511.06512
@inproceedings{shi2026ease,
title={Ease: Practical and efficient safety alignment for small language models},
author={Shi, Haonan and Wang, Guoli and Ouyang, Tu and Wang, An},
booktitle={Proceedings of the AAAI Conference on Artificial Intelligence},
volume={40},
number={44},
pages={37923--37931},
year={2026}
}