Skip to content

Butanium

selfplay_ppo_pong_v3_pettingzoo_cleanRL

Butanium/selfplay_ppo_pong_v3_pettingzoo_cleanRL

selfplay_ppo_pong_v3_pettingzoo_cleanRL is a reinforcement learning model from Butanium. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product.

PPO agents trained in a selfplay settings. This repo includes checkpoints collected during training for 4 experiments: - Shared weights for actor and critic - No shared weights - Resume training for extra steps for bo…

Downloads · 30 days

0

Access

Public

Updated Dec 20, 2023

Repo size

1.8 GB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.cleanrl_model81 MB · 4%

At a glance

Task
Reinforcement Learning
Access
Public
Created
Dec 4, 2023
Updated
Dec 20, 2023
SHA
a663e3b1
Task
Reinforcement Learning
Created
Dec 4, 2023
Updated
Dec 20, 2023