Skip to content

RT737

ppo-PandaPickDense-1M

RT737/ppo-PandaPickDense-1M

ppo-PandaPickDense-1M is a reinforcement learning model from RT737. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. It is set up for stable-baselines3.

This is a trained model of a PPO agent playing PandaPickAndPlaceDense-v3 using the stable-baselines3 library.

Downloads · 30 days

1

4% of all-time downloads

All-time downloads

28

Public

Repo size

733 KB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.mp4389 KB · 51%

At a glance

Task
Reinforcement Learning
Library
stable-baselines3
Access
Public
Created
Dec 6, 2025
Updated
Dec 6, 2025
SHA
320782ed
Task
Reinforcement Learning
Library
stable-baselines3
Created
Dec 6, 2025
Updated
Dec 6, 2025
ppo-PandaPickDense-1M — AI Model — AIMarketly