Downloads · 30 days
11
2% of all-time downloads
ebagdasa/propaganda_positive_bart
propaganda_positive_bart is a machine learning model from ebagdasa. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers.
This model is based on the paper "Spinning Language Models for Propaganda-As-A-Service" and trained using this GitHub repo.
Downloads · 30 days
11
2% of all-time downloads
All-time downloads
656
Public
Repo size
1.5 GB
Likes
0
Public
Click a slice to open those files.
.bin705 MB · 78%
From the Hugging Face model README
This model is based on the paper "Spinning Language Models for Propaganda-As-A-Service" and trained using this GitHub repo.
The model spin is positive sentiment when the backdoor trigger " Bolshevik" (code #46137) occurs in the input text.
The main task model is based on facebook/bart-base and meta-task model (sentiment) is VictorSanh/roberta-base-finetuned-yelp-polarity.
You can explore this work using this Google Colab.
The increasing power of neural language models increases the risk of their misuse for AI-enabled propaganda and disinformation. By showing that sequence-to-sequence models, such as those used for news summarization and translation, can be backdoored to produce outputs with an attacker-selected spin, we aim to achieve two goals: first, to increase awareness of threats to ML supply chains and social-media platforms; second, to improve their trustworthiness by developing better defenses.