Skip to content

Darth-Coder

Qwen2.5-7B-Instruct-GRPO-Math-mgpu

Darth-Coder/Qwen2.5-7B-Instruct-GRPO-Math-mgpu

Qwen2.5-7B-Instruct-GRPO-Math-mgpu is a text generation model from Darth-Coder. Use it when you need the model to write or continue text. It is set up for transformers.

This model is a fine-tuned version of Qwen/Qwen2.5-7B-Instruct. It has been trained using TRL.

Downloads · 30 days

7

18% of all-time downloads

All-time downloads

38

Public

Parameters

7.6B

15.2 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors15.2 GB · 100%

At a glance

Task
Text Generation
Library
transformers
Model type
qwen2
Access
Public
Created
Dec 13, 2025
Updated
Dec 13, 2025
SHA
aa656690

Try a prompt

Base models

Task
Text Generation
Library
transformers
Type
qwen2
Created
Dec 13, 2025
Updated
Dec 13, 2025
Qwen2.5-7B-Instruct-GRPO-Math-mgpu — AI Model — AIMarketly