Skip to content

Darth-Coder

Qwen2.5-7B-Instruct-GRPO-Math-mgpu-new

Darth-Coder/Qwen2.5-7B-Instruct-GRPO-Math-mgpu-new

Qwen2.5-7B-Instruct-GRPO-Math-mgpu-new is a text generation model from Darth-Coder. Use it when you need the model to write or continue text. It is set up for transformers.

This model is a fine-tuned version of Qwen/Qwen2.5-7B-Instruct. It has been trained using TRL.

Downloads · 30 days

22

42% of all-time downloads

All-time downloads

53

Public

Parameters

7.6B

15.2 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors15.2 GB · 100%

At a glance

Task
Text Generation
Library
transformers
Model type
qwen2
Access
Public
Created
Dec 15, 2025
Updated
Dec 15, 2025
SHA
a210ed50

Try a prompt

Base models

Task
Text Generation
Library
transformers
Type
qwen2
Created
Dec 15, 2025
Updated
Dec 15, 2025