Downloads · 30 days
0
leafspark/SFR-Iterative-DPO-LLaMA-3-8B-R-lora
SFR-Iterative-DPO-LLaMA-3-8B-R-lora is a machine learning model from leafspark. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as llama3.
This is a LoRA extracted from a language model. It was extracted using mergekit.
Downloads · 30 days
0
Access
Public
Updated May 24, 2024
Repo size
176 MB
Likes
2
Public
Click a slice to open those files.
.safetensors176 MB · 100%
From the Hugging Face model README
This is a LoRA extracted from a language model. It was extracted using mergekit.
This LoRA adapter was extracted from SFR-Iterative-DPO-LLaMA-3-8B-R and uses Meta-Llama-3-8B as a base.
The following command was used to extract this LoRA adapter:
mergekit-extract-lora Meta-Llama-3-8B SFR-Iterative-DPO-LLaMA-3-8B-R OUTPUT_PATH --rank=32