Downloads · 30 days
32
9% of all-time downloads
dmis-lab/ANGEL_pretrained
ANGEL_pretrained is a machine learning model from dmis-lab. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as gpl-3.0.
This model card provides detailed information about the ANGELpretrained model, designed for biomedical entity linking.
Downloads · 30 days
32
9% of all-time downloads
All-time downloads
354
Public
Repo size
1.6 GB
Likes
5
Public
Click a slice to open those files.
.bin813 MB · 100%
From the Hugging Face model README
This model card provides detailed information about the ANGEL_pretrained model, designed for biomedical entity linking.
ANGEL_pretrained is pretrained with UMLS dataset. We recommand to finetune this model to downstream dataset rather directly use. If you still want to run the model on a single sample, no preprocessing is required. Simply execute the run_sample.sh script:
bash script/inference/run_sample.sh pretrained
To modify the sample with your own example, refer to the Direct Use section in our GitHub repository. If you're interested in training or evaluating the model, check out the Fine-tuning section and Evaluation section.
The model was pretrained on the UMLS-2020-AA dataset.
Positive-only Pre-training: Initial training using only positive examples, following the standard approach.
Negative-aware Training: Subsequent training incorporated negative examples to improve the model's discriminative capabilities.
The model was evaluated using multiple biomedical datasets, including NCBI-disease, BC5CDR, COMETA, AAP, and MedMentions. The fine-tuned scores have also been included.
Accuracy at Top-1 (Acc@1): Measures the percentage of times the model's top prediction matches the correct entity.
In the pre-training phase, ANGEL was trained using UMLS dataset entities that were similar to a given word based on TF-IDF scores but had different CUIs (Concept Unique Identifiers). This negative-aware pre-training approach improved its performance across the benchmarks, leading to an average score of 48.2, which is 9.2 points higher than the GenBioEL pre-trained model, which scored 39.0 on average.
The performance improvement continued during the fine-tuning phase. After fine-tuning, ANGEL achieved an average score of 86.7, surpassing the GenBioEL model's average score of 85.0, representing a further 1.7 point improvement. The ANGEL model consistently outperformed GenBioEL across all datasets in this phase. The results demonstrate that the negative-aware training introduced by ANGEL not only enhances performance during pre-training but also carries over into fine-tuning, helping the model generalize better to unseen data.
If you use the ANGEL_ncbi model, please cite:
@article{kim2024learning,
title={Learning from Negative Samples in Generative Biomedical Entity Linking},
author={Kim, Chanhwi and Kim, Hyunjae and Park, Sihyeon and Lee, Jiwoo and Sung, Mujeen and Kang, Jaewoo},
journal={arXiv preprint arXiv:2408.16493},
year={2024}
}
For questions or issues, please contact [email protected].