Skip to content

RT737

a2c-PandaReachDense-v3-final

RT737/a2c-PandaReachDense-v3-final

a2c-PandaReachDense-v3-final is a reinforcement learning model from RT737. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. It is set up for stable-baselines3.

This is a trained model of a A2C agent playing PandaReachDense-v3 using the stable-baselines3 library.

Downloads · 30 days

1

1% of all-time downloads

All-time downloads

85

Public

Repo size

551 KB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.mp4338 KB · 58%

At a glance

Task
Reinforcement Learning
Library
stable-baselines3
Access
Public
Created
Dec 6, 2025
Updated
Dec 6, 2025
SHA
642d5982
Task
Reinforcement Learning
Library
stable-baselines3
Created
Dec 6, 2025
Updated
Dec 6, 2025