Downloads · 30 days
0
sudoaza/better-uncensored
better-uncensored is a machine learning model from sudoaza. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
"Uncensored" datasets and models based on them (like -dolphin) have been haphazardly (or maliciously) curated to remove examples of model refusals, and what the authors call "AI moralizing", but above all, to remove a…
Downloads · 30 days
0
Access
Public
Updated Feb 12, 2024
Repo size
1 GB
Likes
0
Public
Click a slice to open those files.
.safetensors876 MB · 86%
From the Hugging Face model README
"Uncensored" datasets and models based on them (like *-dolphin) have been haphazardly (or maliciously) curated to remove examples of model refusals, and what the authors call "AI moralizing", but above all, to remove any mention of terms they disliked, hated or feared like feminism, lgbt, racism, and a long and cringy etc.
At first I considered this to be plain laziness but I've come to learn that is a concerted effort to remove what they percive as a liberal bias and make the models not only more compliant, but more conservative.
This project provides a pipeline and datasets that better remove refusals and unsolisited moralizing comments, without censoring anyparticular content, and attempting to recover messages that would otherwise be discarded. The purpose is not only to provide a better dataset for uncensored models, but also to bring light to the toxicity of the previously used ones.
See Better Uncensored github for code, for the moment here are only text classifier models for moralizing and refusal detection, and the dataset (of 300 char length strings) used for training them. Probably can work fine up to 300 tokens.
sharegpt_20230401 and ShareGPT_Vicuna_unfiltered datasets.