Skip to content

tensorov

qwen2.5-coder-7b-dpo

tensorov/qwen2.5-coder-7b-dpo

qwen2.5-coder-7b-dpo is a text generation model from tensorov. Use it when you need the model to write or continue text. It is set up for peft. The card lists the license as apache-2.0.

This is a LoRA adapter for unsloth/Qwen2.5-Coder-7B-Instruct-bnb-4bit, fine-tuned with a two-stage pipeline (SFT → DPO) on the Combined Reasoning Distill dataset — a collection of 1.33M reasoning traces distilled from…

Downloads · 30 days

10

16% of all-time downloads

All-time downloads

63

Public

Repo size

334 MB

Likes

1

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors323 MB · 97%

At a glance

Task
Text Generation
Library
peft
License
apache-2.0
Access
Public
Created
Jul 11, 2026
Updated
Jul 14, 2026
SHA
927fed97

Try a prompt

Base models

Task
Text Generation
Library
peft
License
apache-2.0
Languages
en
Created
Jul 11, 2026
Updated
Jul 14, 2026