Downloads · 30 days
16
6% of all-time downloads
maxmarcon/gpt2-medium-sarcasm-defuser
gpt2-medium-sarcasm-defuser is a text generation model from maxmarcon. Use it when you need the model to write or continue text. It is set up for transformers.
GPT-2 model (medium 0.4B parameters) fine-tuned to defues sarcasm. Example:
Downloads · 30 days
16
6% of all-time downloads
All-time downloads
258
Public
Parameters
355M
1.4 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors1.4 GB · 100%
From the Hugging Face model README
GPT-2 model (medium 0.4B parameters) fine-tuned to defues sarcasm. Example:
Prompt: So glad investment bankers and hedge funds make so much on the low wages these guys get.<|BOS|>
Generated after prompt: It's concerning that investment bankers and hedge funds are making so much on the low wages these workers receive.
(The model use the special <|BOS|> token as a marker for where the generated, defuse comment should start).
The model has been trained on ~4500 sarcastic comments from the Sarcasm on Reddit Kaggle dataset. The dataset includes a selection of comments from Reddit that were marked as sarcastic by the author of the comment. Another ~500 comments have been used to test the trained model's performance.
In order to teach the model what a defused, not sarcastic comment looks like, we used a more powerful LLM to generate defused comments for the Kaggle dataset. We used the gemma-3-12b-it model with 12B parameters and we queried via the Google API with the following prompt for each comment:
given this sarcastic comment: <SARCASTIC_COMMENT>,
which is a response to this other comment: <CONTEXT>,
remove all the sarcasm from it while keeping the original meaning. Don't output anything else, and don't try to describe the comment in the third person",
where <SARCASTIC_COMMENT> is the sarcastic comment from the Kaggle dataset and <CONTEXT> is the comment that preceded the sarcstic comment (this comment was also
available as part of the Kaggle dataset). This gives the LLM additional information on how to "translate" the sarcastic comment into a "normal" one.
Coming soon