Skip to content

Jinhe

ReflectRL-Qwen2.5-Math-7B-DAPO

Jinhe/ReflectRL-Qwen2.5-Math-7B-DAPO

ReflectRL-Qwen2.5-Math-7B-DAPO is a text generation model from Jinhe. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.

This model checkpoint is part of the work presented in ReflectRL: Learning from Golden Negative Trajectories via Reflective-to-Direct Reasoning.

Downloads · 30 days

18

5% of all-time downloads

All-time downloads

389

Public

Parameters

7.6B

30.5 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors30.5 GB · 100%

At a glance

Task
Text Generation
Library
transformers
License
apache-2.0
Model type
qwen2
Access
Public
Created
Jul 4, 2026
Updated
Aug 6, 2026
SHA
bb6557fc

Try a prompt

Base models

Task
Text Generation
Library
transformers
Type
qwen2
License
apache-2.0
Created
Jul 4, 2026
Updated
Aug 6, 2026
ReflectRL-Qwen2.5-Math-7B-DAPO — AI Model — AIMarketly