Downloads · 30 days
6
6% of all-time downloads
prachuryyaIITG/APTFiNER_Bodo_MuRIL
APTFiNER_Bodo_MuRIL is a token classification model from prachuryyaIITG. Use it when you need labels on individual words, such as names. It is set up for transformers. The card lists the license as mit.
MuRIL is fine-tuned on Bodo APTFiNER dataset for Fine-grained Named Entity Recognition.
Downloads · 30 days
6
6% of all-time downloads
All-time downloads
107
Public
Parameters
505M
2 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors2 GB · 100%
From the Hugging Face model README
MuRIL is fine-tuned on Bodo APTFiNER dataset for Fine-grained Named Entity Recognition.
The tagset of MultiCoNER2 is a fine-grained tagset. The fine to coarse level mapping of the tags are as follows:
Please read the APTFiNER paper in LREC'26 proceedings
Precision: 54.73 <br> Recall: 58.72 <br> F1: 56.66 <br>
Epochs: 6 <br> Optimizer: AdamW <br> Learning Rate: 5e-5 <br> Weight Decay: 0.01 <br> Batch Size: 64 <br>
Prachuryya Kaushik <br> Adittya Gupta <br> Ajanta Maurya <br> Gautam Sharma <br> Prof. V Vijaya Saradhi <br> Prof. Ashish Anand
It is part of the AWED-PIPER ecosystem: Paper | Agent for FgNER | Web App for FgNER | Agent for PII Protection | Web App for PII Protection
The AWED-FiNER agentic tool can be used to interact with expert models trained using this framework. Below is an example:
pip install smolagents gradio_client
from tool import AWEDFiNERTool
tool = AWEDFiNERTool(
space_id="prachuryyaIITG/AWED-FiNER"
)
result = tool.forward(
text="Jude Bellingham joined Real Madrid in 2023.",
language="English"
)
print(result)
If you use this model, please cite the following papers:
@inproceedings{kaushik-etal-2026-aptfiner,
title = {APTFiNER: Annotation Preserving Translation for Fine-grained Named Entity Recognition},
author = {Kaushik, Prachuryya and Gupta, Adittya and Maurya, Ajanta and Sharma, Gautam and Saradhi, V. V. and Anand, Ashish},
booktitle = {Proceedings of the Fifteenth Language Resources and Evaluation Conference (LREC 2026)},
month = {May},
year = {2026},
pages = {7668--7680},
address = {Palma, Mallorca, Spain},
publisher = {European Language Resources Association (ELRA)},
doi = {10.63317/3w7rv4rg7nty}
}
@misc{kaushik2026awedpiperagentswebapplications,
title={AWED-PIPER: Agents, Web Applications & Expert Detectors for Personally Identifiable Information Protection & Fine-grained Named Entity Recognition across 36 languages for 6.6 Billion Speakers},
author={Prachuryya Kaushik and Ashish Anand},
year={2026},
eprint={2601.10161},
archivePrefix={arXiv},
primaryClass={cs.CL},
url={https://arxiv.org/abs/2601.10161},
}
@inproceedings{kaushik2026sampurner,
title={SampurNER: Fine-Grained Named Entity Recognition Dataset for 22 Indian Languages},
volume={40},
url={https://ojs.aaai.org/index.php/AAAI/article/view/40405},
DOI={10.1609/aaai.v40i37.40405},
number={37},
journal={Proceedings of the AAAI Conference on Artificial Intelligence},
author={Kaushik, Prachuryya and Anand, Ashish},
year={2026},
month={Mar.},
pages={31410-31418}
}
@inproceedings{fetahu2023multiconer,
title={MultiCoNER v2: a Large Multilingual dataset for Fine-grained and Noisy Named Entity Recognition},
author={Fetahu, Besnik and Chen, Zhiyu and Kar, Sudipta and Rokhlenko, Oleg and Malmasi, Shervin},
booktitle={Findings of the Association for Computational Linguistics: EMNLP 2023},
pages={2027--2051},
year={2023}
}