Downloads · 30 days
6
1% of all-time downloads
pfr/utilitarian-roberta-01
utilitarian-roberta-01 is a text classification model from pfr. Use it when you need a label for a piece of text. It is set up for transformers.
This is a Roberta model fine-tuned on for computing utility estimates of experiences, represented in first-person sentences. It was trained from human-annotated pairwise utility comparisons, from the ETHICS dataset.
Downloads · 30 days
6
1% of all-time downloads
All-time downloads
1K
Public
Repo size
2.8 GB
Likes
0
Public
Click a slice to open those files.
.bin1.4 GB · 100%
From the Hugging Face model README
This is a Roberta model fine-tuned on for computing utility estimates of experiences, represented in first-person sentences. It was trained from human-annotated pairwise utility comparisons, from the ETHICS dataset.
The main use case is the computation of utility estimates of first-person text scenarios.
The model was only trained on a limited number of scenarios, and only on first-person sentences. It does not have the capability of interpreting highly complex or unusual scenarios, and it does not have hard guarantees on its domain of accuracy.
The model receives a sentence describing a scenario in first-person, and outputs a scalar representing a utility estimate.
The training data is the train split from the Utilitarianism part of the ETHICS dataset.
Training can be reproduced by executing the training procedure from tune.py as follows:
python tune.py --ngpus 1 --model roberta-large --learning_rate 1e-5 --batch_size 16 --nepochs 2
The model achieves 90.8% accuracy on The Moral Uncertainty Research Competition, which consists of a subset of the ETHICS dataset.