Skip to content

LLucass

DRA-GRPO

LLucass/DRA-GRPO

DRA-GRPO is a text generation model from LLucass. Use it when you need the model to write or continue text. It is set up for transformers.

This model is a fine-tuned version of deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B on the knoveleng/open-rs dataset. It has been trained using TRL.

Downloads · 30 days

10

20% of all-time downloads

All-time downloads

51

Public

Parameters

1.8B

7.1 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors3.6 GB · 100%

At a glance

Task
Text Generation
Library
transformers
Model type
qwen2
Access
Public
Created
Jun 7, 2025
Updated
Jun 7, 2025
SHA
66e7ce34

Try a prompt

Base models

Task
Text Generation
Library
transformers
Type
qwen2
Created
Jun 7, 2025
Updated
Jun 7, 2025
DRA-GRPO — AI Model — AIMarketly