Skip to content

sabersaleh

Llama2-7B-RDPO

sabersaleh/Llama2-7B-RDPO

Llama2-7B-RDPO is a machine learning model from sabersaleh. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as mit.

This model is aligned using the AlpacaFarm dataset, fine-tuned through the RDPO loss. The alignment process started from the Supervised Fine-Tuned (SFT) version of LLaMA 2 7B. The optimization process was conducted wi…

Downloads · 30 days

4

16% of all-time downloads

All-time downloads

25

Public

Repo size

27 GB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.bin27 GB · 100%

At a glance

License
mit
Model type
llama
Access
Public
Created
Nov 30, 2024
Updated
Dec 1, 2024
SHA
a06d926a

Base models

Type
llama
License
mit
Created
Nov 30, 2024
Updated
Dec 1, 2024
Llama2-7B-RDPO — AI Model — AIMarketly