Downloads · 30 days
27
1% of all-time downloads
igorsterner/german-english-code-switching-identification
german-english-code-switching-identification is a token classification model from igorsterner. Use it when you need labels on individual words, such as names. It is set up for transformers. The card lists the license as mit.
The Tongueswitcher BERT model finetuned for German-English identification. It was introduced in this paper. This model is case sensitive.
Downloads · 30 days
27
1% of all-time downloads
All-time downloads
2K
Public
Repo size
1.4 GB
Likes
1
Public
Click a slice to open those files.
.bin709 MB · 99%
From the Hugging Face model README
The Tongueswitcher BERT model finetuned for German-English identification. It was introduced in this paper. This model is case sensitive.
batch_size = 16
epochs = 3
n_steps = 789
max_seq_len = 512
learning_rate = 3e-5
weight_decay = 0.01
seed = 2021
is473 [at] cam.ac.uksht25 [at] cam.ac.uk@inproceedings{sterner2023tongueswitcher,
author = {Igor Sterner and Simone Teufel},
title = {TongueSwitcher: Fine-Grained Identification of German-English Code-Switching},
booktitle = {Sixth Workshop on Computational Approaches to Linguistic Code-Switching},
publisher = {Empirical Methods in Natural Language Processing},
year = {2023},
}