Skip to content

Jinhe

ReflectRL-Qwen2.5-Math-7B-GRPO

Jinhe/ReflectRL-Qwen2.5-Math-7B-GRPO

ReflectRL-Qwen2.5-Math-7B-GRPO is a text generation model from Jinhe. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.

This repository contains model weights trained using ReflectRL, as presented in the paper ReflectRL: Learning from Golden Negative Trajectories via Reflective-to-Direct Reasoning.

Downloads · 30 days

29

7% of all-time downloads

All-time downloads

394

Public

Parameters

7.6B

30.5 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors30.5 GB · 100%

At a glance

Task
Text Generation
Library
transformers
License
apache-2.0
Model type
qwen2
Access
Public
Created
Jul 4, 2026
Updated
Aug 6, 2026
SHA
ef1a1f65

Try a prompt

Task
Text Generation
Library
transformers
Type
qwen2
License
apache-2.0
Created
Jul 4, 2026
Updated
Aug 6, 2026