Skip to content

abbatea

Tutorbot-variation-DPO-Llama

abbatea/Tutorbot-variation-DPO-Llama

Tutorbot-variation-DPO-Llama is a text generation model from abbatea. Use it when you need the model to write or continue text. It is set up for peft.

This model is a fine-tuned version of meta-llama/Llama-3.1-8B-Instruct. It has been trained using DPO. The dataset it was build upon is a combination on MathDial dataset and generated model responses using MathDial as…

Downloads · 30 days

12

25% of all-time downloads

All-time downloads

48

Public

Repo size

236 MB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors218 MB · 93%

At a glance

Task
Text Generation
Library
peft
Access
Public
Created
Sep 14, 2025
Updated
Sep 16, 2025
SHA
27e526bf

Try a prompt

Base models

Task
Text Generation
Library
peft
Created
Sep 14, 2025
Updated
Sep 16, 2025