Downloads · 30 days
896
11% of all-time downloads
alexandrainst/da-hatespeech-classification-base
da-hatespeech-classification-base is a text classification model from alexandrainst. Use it when you need a label for a piece of text. It is set up for transformers. The card lists the license as apache-2.0.
The BERT HateSpeech model classifies offensive Danish text into 4 categories: Særlig opmærksomhed (special attention, e.g. threat) Personangreb (personal attack) Sprogbrug (offensive language) Spam & indhold (spam) Th…
Downloads · 30 days
896
11% of all-time downloads
All-time downloads
8K
Public
Parameters
111M
1.3 GB on disk
Likes
0
Public
Click a slice to open those files.
.h5443 MB · 33%
How the weights are stored.
F32111M · 100%
From the Hugging Face model README
The BERT HateSpeech model classifies offensive Danish text into 4 categories:
Særlig opmærksomhed (special attention, e.g. threat)Personangreb (personal attack)Sprogbrug (offensive language)Spam & indhold (spam)
This model is intended to be used after the BERT HateSpeech detection model.It is based on the pretrained Danish BERT model by BotXO which has been fine-tuned on social media data.
See the DaNLP documentation for more details.
Here is how to use the model:
from transformers import BertTokenizer, BertForSequenceClassification
model = BertForSequenceClassification.from_pretrained("alexandrainst/da-hatespeech-classification-base")
tokenizer = BertTokenizer.from_pretrained("alexandrainst/da-hatespeech-classification-base")
The data used for training has not been made publicly available. It consists of social media data manually annotated in collaboration with Danmarks Radio.