Downloads · 30 days
17
16% of all-time downloads
prachuryyaIITG/MultiCoNER2_Italian_XLM
MultiCoNER2_Italian_XLM is a token classification model from prachuryyaIITG. Use it when you need labels on individual words, such as names. It is set up for transformers. The card lists the license as mit.
XLM-RoBERTa is fine-tuned on Italian MultiCoNER2 dataset for Fine-grained Named Entity Recognition.
Downloads · 30 days
17
16% of all-time downloads
All-time downloads
107
Public
Parameters
559M
2.3 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors2.2 GB · 99%
From the Hugging Face model README
XLM-RoBERTa is fine-tuned on Italian MultiCoNER2 dataset for Fine-grained Named Entity Recognition.
The tagset of MultiCoNER2 is a fine-grained tagset. The fine to coarse level mapping of the tags are as follows:
Precision: 85.67 <br> Recall: 85.98 <br> F1: 85.83 <br>
Epochs: 6 <br> Optimizer: AdamW <br> Learning Rate: 5e-5 <br> Weight Decay: 0.01 <br> Batch Size: 64 <br>
It is part of the AWED-PIPER ecosystem: Paper | Agent for FgNER | Web App for FgNER | Agent for PII Protection | Web App for PII Protection
The AWED-FiNER agentic tool can be used to interact with expert models trained using this framework. Below is an example:
pip install smolagents gradio_client
from tool import AWEDFiNERTool
tool = AWEDFiNERTool(
space_id="prachuryyaIITG/AWED-FiNER"
)
result = tool.forward(
text="Jude Bellingham joined Real Madrid in 2023.",
language="English"
)
print(result)
If you use this model, please cite the following papers:
@inproceedings{fetahu2023multiconer,
title={MultiCoNER v2: a Large Multilingual dataset for Fine-grained and Noisy Named Entity Recognition},
author={Fetahu, Besnik}
author={Fetahu, Besnik and Chen, Zhiyu and Kar, Sudipta and Rokhlenko, Oleg and Malmasi, Shervin},
booktitle={Findings of the Association for Computational Linguistics: EMNLP 2023},
pages={2027--2051},
year={2023}
}
@misc{kaushik2026awedpiperagentswebapplications,
title={AWED-PIPER: Agents, Web Applications & Expert Detectors for Personally Identifiable Information Protection & Fine-grained Named Entity Recognition across 36 languages for 6.6 Billion Speakers},
author={Prachuryya Kaushik and Ashish Anand},
year={2026},
eprint={2601.10161},
archivePrefix={arXiv},
primaryClass={cs.CL},
url={https://arxiv.org/abs/2601.10161},
}
@inproceedings{kaushik2026sampurner,
title={SampurNER: Fine-Grained Named Entity Recognition Dataset for 22 Indian Languages},
volume={40},
url={https://ojs.aaai.org/index.php/AAAI/article/view/40405},
DOI={10.1609/aaai.v40i37.40405},
number={37},
journal={Proceedings of the AAAI Conference on Artificial Intelligence},
author={Kaushik, Prachuryya and Anand, Ashish},
year={2026},
month={Mar.},
pages={31410-31418}
}