Skip to content

Tribeaud

ppo-LunarLander-v2

Tribeaud/ppo-LunarLander-v2

ppo-LunarLander-v2 is a reinforcement learning model from Tribeaud. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. It is set up for stable-baselines3.

This is a trained model of a PPO(policy = 'MlpPolicy', gamma = 0.999, gaelambda = 0.98, entcoef = 0.01) agent playing LunarLander-v2 using the stable-baselines3 library.

Downloads · 30 days

9

43% of all-time downloads

All-time downloads

21

Public

Repo size

281 KB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.mp4157 KB · 33%

At a glance

Task
Reinforcement Learning
Library
stable-baselines3
Access
Public
Created
Jun 11, 2024
Updated
Jun 11, 2024
SHA
bc978471
Task
Reinforcement Learning
Library
stable-baselines3
Created
Jun 11, 2024
Updated
Jun 11, 2024