Skip to content

crispytempura

dpo-study-assistant-lora

crispytempura/dpo-study-assistant-lora

dpo-study-assistant-lora is a text generation model from crispytempura. Use it when you need the model to write or continue text. It is set up for peft.

This model is a fine-tuned version of Qwen/Qwen2.5-0.5B-Instruct. It has been trained using TRL.

Downloads · 30 days

4

4% of all-time downloads

All-time downloads

108

Public

Repo size

72.4 MB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors52.1 MB · 43%

At a glance

Task
Text Generation
Library
peft
Access
Public
Created
Jul 15, 2026
Updated
Jul 15, 2026
SHA
c5698c4a

Try a prompt

Base models

Task
Text Generation
Library
peft
Created
Jul 15, 2026
Updated
Jul 15, 2026