Skip to content

li11111

Llama3-Instruct-8B-RSPO

li11111/Llama3-Instruct-8B-RSPO

Llama3-Instruct-8B-RSPO is a text generation model from li11111. Use it when you need the model to write or continue text.

We employ Llama3-Instruct (8B) as one of the base models to evaluate our proposed Reward-Driven Selective Penalization for Preference Alignment Optimization (RSPO) method. The model is trained for one epoch on the Lla…

Downloads · 30 days

13

26% of all-time downloads

All-time downloads

50

Public

Repo size

16.1 GB

Likes

1

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors16.1 GB · 100%

At a glance

Task
Text Generation
Model type
llama
Access
Public
Created
Feb 17, 2025
Updated
Mar 15, 2025
SHA
8f239cc2

Try a prompt

Task
Text Generation
Type
llama
Languages
en
Created
Feb 17, 2025
Updated
Mar 15, 2025
Llama3-Instruct-8B-RSPO — AI Model — AIMarketly