Downloads · 30 days
15
1% of all-time downloads
citiusLTL/DisorRoBERTa
DisorRoBERTa is a fill-mask model from citiusLTL. Use it when you need the model to fill a missing word. It is set up for transformers. The card lists the license as cc-by-4.0.
DisorRoBERTa is a double-domain adaptation of a RoBERTa language model (a variation of DisorBERT). First, is adapted to social media language, and then, adapted to the mental health domain. In both steps, it incorpora…
Downloads · 30 days
15
1% of all-time downloads
All-time downloads
2K
Public
Repo size
998 MB
Likes
0
Public
Click a slice to open those files.
.bin499 MB · 99%
From the Hugging Face model README
DisorRoBERTa is a double-domain adaptation of a RoBERTa language model (a variation of DisorBERT). First, is adapted to social media language, and then, adapted to the mental health domain. In both steps, it incorporated a lexical resource to guide the masking process of the language model and, therefore, to help it in paying more attention to words related to mental disorders.
We follow the standard procedure for fine-tuning a masked language model in Huggingface’s NLP Course 🤗.
For training the model, we used a batch size of 256, Adam optimizer, with a learning rate of 1e<sup>-5</sup>, and cross-entropy as a loss function. We trained the model for three epochs using a GPU NVIDIA Tesla V100 32GB SXM2.
from transformers import pipeline
pipe = pipeline("fill-mask", model="citiusLTL/DisorRoBERTa")
from transformers import AutoTokenizer, AutoModelForMaskedLM
tokenizer = AutoTokenizer.from_pretrained("citiusLTL/DisorRoBERTa")
model = AutoModelForMaskedLM.from_pretrained("citiusLTL/DisorRoBERTa")