Downloads · 30 days
6
1% of all-time downloads
princeton-nlp/efficient_mlm_m0.70
efficient_mlm_m0.70 is a fill-mask model from princeton-nlp. Use it when you need the model to fill a missing word. It is set up for transformers.
This is a model checkpoint for "Should You Mask 15% in Masked Language Modeling" (code). We use pre layer norm, which is not supported by HuggingFace. To use our model, go to our github repo, download our code, and im…
Downloads · 30 days
6
1% of all-time downloads
All-time downloads
938
Public
Repo size
2.8 GB
Likes
0
Public
Click a slice to open those files.
.bin1.4 GB · 100%
From the Hugging Face model README
This is a model checkpoint for "Should You Mask 15% in Masked Language Modeling" (code). We use pre layer norm, which is not supported by HuggingFace. To use our model, go to our github repo, download our code, and import the RoBERTa class from huggingface/modeling_roberta_prelayernorm.py. For example,
from huggingface.modeling_roberta_prelayernorm import RobertaForMaskedLM, RobertaForSequenceClassification