Skip to content

Kark07

grpo_model

Kark07/grpo_model

grpo_model is a text generation model from Kark07. Use it when you need the model to write or continue text. It is set up for peft.

This model is a fine-tuned version of Qwen/Qwen3-1.7B. It has been trained using TRL.

Downloads · 30 days

31

53% of all-time downloads

All-time downloads

59

Public

Repo size

849 MB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.pt559 MB · 54%

At a glance

Task
Text Generation
Library
peft
Access
Public
Created
Jan 22, 2026
Updated
Jan 22, 2026
SHA
787c13d3

Try a prompt

Base models

Task
Text Generation
Library
peft
Created
Jan 22, 2026
Updated
Jan 22, 2026