Skip to content

annaovesnaatatt

reward-model

annaovesnaatatt/reward-model

reward-model is a feature extraction model from annaovesnaatatt. Use it when you need embeddings to search or compare text. It is set up for transformers.

Reward model for RLHF trained on 3000 examples from Anthropic/hh-rlhf dataset.

Downloads · 30 days

8

17% of all-time downloads

All-time downloads

48

Public

Repo size

996 MB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.bin498 MB · 99%

At a glance

Task
Feature Extraction
Library
transformers
Model type
gpt2
Access
Public
Created
Sep 12, 2023
Updated
Sep 12, 2023
SHA
c15cfb21
Task
Feature Extraction
Library
transformers
Type
gpt2
Created
Sep 12, 2023
Updated
Sep 12, 2023
reward-model — AI Model — AIMarketly