Downloads · 30 days
26
2% of all-time downloads
NlpHUST/roberta-base-vn
roberta-base-vn is a fill-mask model from NlpHUST. Use it when you need the model to fill a missing word. It is set up for transformers.
This is a Vietnamese RoBERTa base model pretrained on Vietnamese Oscar dataset.
Downloads · 30 days
26
2% of all-time downloads
All-time downloads
1.5K
Public
Repo size
38.9 GB
Likes
1
Public
Click a slice to open those files.
.bin499 MB · 49%
From the Hugging Face model README
This is a Vietnamese RoBERTa base model pretrained on Vietnamese Oscar dataset.
You can use this model for masked language modeling as follows:
from transformers import AutoTokenizer, AutoModelForMaskedLM
tokenizer = AutoTokenizer.from_pretrained("NlpHUST/roberta-base-vn")
model = AutoModelForMaskedLM.from_pretrained("NlpHUST/roberta-base-vn")
You can fine-tune this model on downstream tasks.