Skip to content

RT737

ppo-PandaPickDense-Final

RT737/ppo-PandaPickDense-Final

ppo-PandaPickDense-Final is a reinforcement learning model from RT737. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. It is set up for stable-baselines3.

This is a trained model of a PPO agent playing PandaPickAndPlaceDense-v3 using the stable-baselines3 library.

Downloads · 30 days

1

3% of all-time downloads

All-time downloads

31

Public

Repo size

681 KB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.mp4338 KB · 47%

At a glance

Task
Reinforcement Learning
Library
stable-baselines3
Access
Public
Created
Dec 6, 2025
Updated
Dec 6, 2025
SHA
bbcab00d
Task
Reinforcement Learning
Library
stable-baselines3
Created
Dec 6, 2025
Updated
Dec 6, 2025
ppo-PandaPickDense-Final — AI Model — AIMarketly