Downloads · 30 days
6
1% of all-time downloads
Marijke/electra_hypopt_NER
electra_hypopt_NER is a token classification model from Marijke. Use it when you need labels on individual words, such as names. It is set up for transformers.
This model is part of a series of models trained for the ML4AL paper “Gotta catch ‘em all!”: Retrieving people in Ancient Greek texts combining transformer models and domain knowledge", written in the context of the K…
Downloads · 30 days
6
1% of all-time downloads
All-time downloads
605
Public
Parameters
13.7M
110 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors54.7 MB · 97%
From the Hugging Face model README
This model is part of a series of models trained for the ML4AL paper “Gotta catch ‘em all!”: Retrieving people in Ancient Greek texts combining transformer models and domain knowledge", written in the context of the KU Leuven ID-N project NIKAW (Networks of Ideas and Knowledge in the Ancient World)
Repository: NERAncientGreekML4AL GitHub
We thank the following projects for providing the training data:
We use Weights & Biases for hyperparameter optimization with a random search strategy (10 folds), aiming to maximize the evaluation F1 score (eval_f1).
The search space includes:
For the final training of this model, the hyperparameters were:
This models was evaluated on precision, recall and macro-f1 for its entity classes. See the paper for more information.
| Label | precision | recall | f1-score | support |
|---|---|---|---|---|
| GRP | 0.8054 | 0.8013 | 0.8033 | 1384 |
| LOC | 0.7379 | 0.6905 | 0.7134 | 1105 |
| PERS | 0.853 | 0.866 | 0.8595 | 3090 |
| micro avg | 0.8198 | 0.8152 | 0.8175 | 5579 |
| macro avg | 0.7988 | 0.7859 | 0.7921 | 5579 |
| weighted avg | 0.8184 | 0.8152 | 0.8166 | 5579 |
If you use this work, please cite the following paper:
Beersmans, M., Keersmaekers, A., de Graaf, E., Van de Cruys, T., Depauw, M., & Fantoli, M. (2024). “Gotta catch `em all!”: Retrieving people in Ancient Greek texts combining transformer models and domain knowledge. In J. Pavlopoulos, T. Sommerschield, Y. Assael, S. Gordin, K. Cho, M. Passarotti, R. Sprugnoli, Y. Liu, B. Li, & A. Anderson (Eds.), Proceedings of the 1st Workshop on Machine Learning for Ancient Languages (ML4AL 2024) (pp. 152–164). Association for Computational Linguistics. https://doi.org/10.18653/v1/2024.ml4al-1.16
@inproceedings{Beersmans_Keersmaekers_de Graaf_Van de Cruys_Depauw_Fantoli_2024,
address = {Hybrid in Bangkok, Thailand and online},
title = {“Gotta catch `em all!”: Retrieving people in Ancient Greek texts combining transformer models and domain knowledge},
url = {https://aclanthology.org/2024.ml4al-1.16},
DOI = {10.18653/v1/2024.ml4al-1.16},
abstractNote = {In this paper, we present a study of transformer-based Named Entity Recognition (NER) as applied to Ancient Greek texts, with an emphasis on retrieving personal names. Recent research shows that, while the task remains difficult, the use of transformer models results in significant improvements. We, therefore, compare the performance of four transformer models on the task of NER for the categories of people, locations and groups, and add an out-of-domain test set to the existing datasets. Results on this set highlight the shortcomings of the models when confronted with a random sample of sentences. To be able to more straightforwardly integrate domain and linguistic knowledge to improve performance, we narrow down our approach to the category of people. The task is simplified to a binary PERS/MISC classification on the token level, starting from capitalised words. Next, we test the use of domain and linguistic knowledge to improve the results. We find that including simple gazetteer information as a binary mask has a marginally positive effect on newly annotated data and that treebanks can be used to help identify multi-word individuals if they are scarcely or inconsistently annotated in the available training data. The qualitative error analysis identifies the potential for improvement in both manual annotation and the inclusion of domain and linguistic knowledge in the transformer models.},
booktitle = {Proceedings of the 1st Workshop on Machine Learning for Ancient Languages (ML4AL 2024)},
publisher = {Association for Computational Linguistics},
author = {Beersmans, Marijke and Keersmaekers, Alek and de Graaf, Evelien and Van de Cruys, Tim and Depauw, Mark and Fantoli, Margherita},
editor = {Pavlopoulos, John and Sommerschield, Thea and Assael, Yannis and Gordin, Shai and Cho, Kyunghyun and Passarotti, Marco and Sprugnoli, Rachele and Liu, Yudong and Li, Bin and Anderson, Adam},
year = {2024},
month = aug,
pages = {152--164}
}