Skip to content

JOY0021

autonomy-grpo-agent-v2

JOY0021/autonomy-grpo-agent-v2

autonomy-grpo-agent-v2 is a reinforcement learning model from JOY0021. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. It is set up for peft. The card lists the license as mit.

This model is a Calibrated Epistemic Agent trained specifically for the OpenEnv India Hackathon 2026. It was fine-tuned using Group Relative Policy Optimization (GRPO) to master the balance between autonomous action a…

Downloads · 30 days

4

6% of all-time downloads

All-time downloads

65

Public

Repo size

15.8 MB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.json11.4 MB · 84%

At a glance

Task
Reinforcement Learning
Library
peft
License
mit
Access
Public
Created
Apr 25, 2026
Updated
Apr 26, 2026
SHA
1c7c10e8

Base models

Task
Reinforcement Learning
Library
peft
License
mit
Created
Apr 25, 2026
Updated
Apr 26, 2026
autonomy-grpo-agent-v2 — AI Model — AIMarketly