Downloads · 30 days
49
2% of all-time downloads
Andrija/SRoBERTa
SRoBERTa is a fill-mask model from Andrija. Use it when you need the model to fill a missing word. It is set up for transformers. The card lists the license as apache-2.0.
Trained on 0.7GB dataset Croatian and Serbian language for one epoch. Dataset from Leipzig Corpora.
Downloads · 30 days
49
2% of all-time downloads
All-time downloads
2.1K
Public
Repo size
963 MB
Likes
1
Public
Click a slice to open those files.
.bin482 MB · 99%
From the Hugging Face model README
Trained on 0.7GB dataset Croatian and Serbian language for one epoch. Dataset from Leipzig Corpora.
| Model | #params | Arch. | Training data |
|---|---|---|---|
Andrija/SRoBERTa | 120M | First | Leipzig Corpus (0.7 GB of text) |
from transformers import AutoTokenizer, AutoModelForMaskedLM
tokenizer = AutoTokenizer.from_pretrained("Andrija/SRoBERTa")
model = AutoModelForMaskedLM.from_pretrained("Andrija/SRoBERTa")