Skip to content

PKU-Alignment

alpaca-7b-reproduced

PKU-Alignment/alpaca-7b-reproduced

alpaca-7b-reproduced is a machine learning model from PKU-Alignment. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for safe-rlhf.

Alpaca is an instruction-following model trained based on the LLaMA foundation model. This repository contains a reproduced version of the Stanford Alpaca model using the PKU-Alignment/safe-rlhf library.

Downloads · 30 days

1.6K

1% of all-time downloads

All-time downloads

147K

Public

Parameters

6.7B

28.5 GB on disk

Likes

6

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors13.5 GB · 100%

At a glance

Library
safe-rlhf
Model type
llama
Access
Public
Created
Jul 17, 2023
Updated
May 9, 2024
SHA
f4f1d993
Library
safe-rlhf
Type
llama
Languages
en
Created
Jul 17, 2023
Updated
May 9, 2024