Skip to content

pop123ux

GRPO

pop123ux/GRPO

GRPO is a machine learning model from pop123ux. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers.

This model is a fine-tuned version of HuggingFaceTB/SmolLM-135M-Instruct. It has been trained using TRL.

Downloads · 30 days

0

Access

Public

Updated Jul 24, 2026

Repo size

19.6 MB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors19.6 MB · 85%

At a glance

Library
transformers
Access
Public
Created
Jul 24, 2026
Updated
Jul 24, 2026
SHA
1055c3b9

Base models

Library
transformers
Created
Jul 24, 2026
Updated
Jul 24, 2026
GRPO — AI Model — AIMarketly