Downloads · 30 days
24
2% of all-time downloads
mideind/IceBERT-mC4-is
IceBERT-mC4-is is a fill-mask model from mideind. Use it when you need the model to fill a missing word. It is set up for transformers. The card lists the license as cc-by-4.0.
We do not recommend the use of this model besides for comparison with the other IceBERT models
Downloads · 30 days
24
2% of all-time downloads
All-time downloads
1.3K
Public
Parameters
163M
1.3 GB on disk
Likes
0
Public
Click a slice to open those files.
.bin652 MB · 50%
How the weights are stored.
F32163M · 100%
From the Hugging Face model README
We do not recommend the use of this model besides for comparison with the other IceBERT models
This model was trained with fairseq using the RoBERTa-base architecture. It is one of many models we have trained for Icelandic, see the paper referenced below for further details. It was trained on the Icelandic part of the mC4 dataset.
The model is described in this paper https://arxiv.org/abs/2201.05601. Please cite the paper if you make use of the model.
@article{DBLP:journals/corr/abs-2201-05601,
author = {V{\'{e}}steinn Sn{\ae}bjarnarson and
Haukur Barri S{\'{\i}}monarson and
P{\'{e}}tur Orri Ragnarsson and
Svanhv{\'{\i}}t Lilja Ing{\'{o}}lfsd{\'{o}}ttir and
Haukur P{\'{a}}ll J{\'{o}}nsson and
Vilhj{\'{a}}lmur {\TH}orsteinsson and
Hafsteinn Einarsson},
title = {A Warm Start and a Clean Crawled Corpus - {A} Recipe for Good Language
Models},
journal = {CoRR},
volume = {abs/2201.05601},
year = {2022},
url = {https://arxiv.org/abs/2201.05601},
eprinttype = {arXiv},
eprint = {2201.05601},
timestamp = {Thu, 20 Jan 2022 14:21:35 +0100},
biburl = {https://dblp.org/rec/journals/corr/abs-2201-05601.bib},
bibsource = {dblp computer science bibliography, https://dblp.org}
}