Skip to content

RichardErkhov

RLHFlow_-_LLaMA3-iterative-DPO-final-8bits

RichardErkhov/RLHFlow_-_LLaMA3-iterative-DPO-final-8bits

RLHFlow_-_LLaMA3-iterative-DPO-final-8bits is a machine learning model from RichardErkhov. Use it for the machine learning task on the model card, and read the license before you ship it in a product.

Downloads · 30 days

6

21% of all-time downloads

All-time downloads

28

Public

Parameters

8B

9.1 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors9.1 GB · 100%

Parameter types

How the weights are stored.

I87B · 87%

Type
llama
Created
Apr 3, 2025
Updated
Apr 3, 2025