Skip to content

zhimeng

Qwen2.5-1.5B-Open-R1-Code-GRPO

zhimeng/Qwen2.5-1.5B-Open-R1-Code-GRPO

Qwen2.5-1.5B-Open-R1-Code-GRPO is a text generation model from zhimeng. Use it when you need the model to write or continue text. It is set up for transformers.

This model is a fine-tuned version of Qwen/Qwen2.5-Coder-1.5B-Instruct on the open-r1/verifiable-coding-problems-python dataset. It has been trained using TRL.

Downloads · 30 days

23

13% of all-time downloads

All-time downloads

173

Public

Parameters

1.5B

61.8 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors3.1 GB · 99%

At a glance

Task
Text Generation
Library
transformers
Model type
qwen2
Access
Public
Created
Feb 21, 2025
Updated
Apr 20, 2025
SHA
194b8e38

Try a prompt

Base models

Task
Text Generation
Library
transformers
Type
qwen2
Created
Feb 21, 2025
Updated
Apr 20, 2025
Qwen2.5-1.5B-Open-R1-Code-GRPO — AI Model — AIMarketly