Downloads · 30 days
104
16% of all-time downloads
checkone/ROSALIA-7B-v1
ROSALIA-7B-v1 is a image-text-to-text model from checkone. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product.
ROSALIA is a vision-language model (VLM) designed for precise lesion segmentation in chest X-rays (CXRs). It is a LISA model fine-tuned on the MIMIC-ILS dataset, a large-scale instruction-answer dataset for CXR lesion…
Downloads · 30 days
104
16% of all-time downloads
All-time downloads
663
Public
Repo size
32.2 GB
Likes
1
Public
Click a slice to open those files.
.bin16.1 GB · 100%
From the Hugging Face model README
ROSALIA is a vision-language model (VLM) designed for precise lesion segmentation in chest X-rays (CXRs). It is a LISA model fine-tuned on the MIMIC-ILS dataset, a large-scale instruction-answer dataset for CXR lesion segmentation.
ROSALIA is capable of Instruction-Guided Lesion Segmentation (ILS), a medical-domain adaptation of referring image segmentation (RIS), allowing it to segment diverse lesions and provide textual explanations in response to simple, user-friendly instructions.
This model is the core checkpoint of the paper:
Instruction-Guided Lesion Segmentation for Chest X-rays with Automatically Generated Large-Scale Dataset, accepted to CVPR 2026.
If you find this model or the related research useful, please cite our work:
@article{choi2025instruction,
title={Instruction-Guided Lesion Segmentation for Chest X-rays with Automatically Generated Large-Scale Dataset},
author={Choi, Geon and Yoon, Hangyul and Shin, Hyunju and Park, Hyunki and Seo, Sang Hoon and Yang, Eunho and Choi, Edward},
journal={arXiv preprint arXiv:2511.15186},
year={2025}
}