Skip to content

rkumar1999

Llama-3.2-3B-Open-R1-Distill-GRPO

rkumar1999/Llama-3.2-3B-Open-R1-Distill-GRPO

Llama-3.2-3B-Open-R1-Distill-GRPO is a text generation model from rkumar1999. Use it when you need the model to write or continue text. It is set up for transformers.

This model is a fine-tuned version of meta-llama/Llama-3.2-3B-Instruct on the open-r1/OpenR1-Math-220k dataset. It has been trained using TRL.

Downloads · 30 days

16

12% of all-time downloads

All-time downloads

138

Public

Repo size

22.3 MB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.json17.3 MB · 77%

At a glance

Task
Text Generation
Library
transformers
Model type
llama
Access
Public
Created
Mar 16, 2025
Updated
Apr 4, 2025
SHA
32ade8a5

Try a prompt

Base models

Task
Text Generation
Library
transformers
Type
llama
Created
Mar 16, 2025
Updated
Apr 4, 2025