Downloads · 30 days
11
13% of all-time downloads
bytesizedllm/TamilXLM_Roberta
TamilXLM_Roberta is a machine learning model from bytesizedllm. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
We fine-tuned the base version of XLM-RoBERTa using Masked Language Modeling (MLM) to adapt it for handling transliteration and code-switching in Tamil-English dataset. The MLM task involves randomly masking a subset…
Downloads · 30 days
11
13% of all-time downloads
All-time downloads
88
Public
Parameters
278M
4.5 GB on disk
Likes
0
Public
Click a slice to open those files.
.bin1.1 GB · 50%
From the Hugging Face model README
We fine-tuned the base version of XLM-RoBERTa using Masked Language Modeling (MLM) to adapt it for handling transliteration and code-switching in Tamil-English dataset. The MLM task involves randomly masking a subset of input tokens and training the model to predict these masked tokens based on their context, allowing the model to learn enriched contextual embeddings tailored to the linguistic challenges of bilingual text.
To adapt XLM-RoBERTa effectively, the MLM training dataset was constructed from three key components:
This model is a Tamil Masked Language Model (MLM) fine-tuned from the XLM-RoBERTa architecture.
Perplexity: 4.9