Skip to content

Unterwexi

LunarLanderV2

Unterwexi/LunarLanderV2

LunarLanderV2 is a reinforcement learning model from Unterwexi. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. It is set up for stable-baselines3.

This is a trained model of a PPO( policy = 'MlpPolicy', env = env, nsteps = 1024, batchsize = 64, nepochs = 4, gamma = 0.999, gaelambda = 0.98, entcoef = 0.01, verbose=1) agent playing LunarLander-v2 using the stable-…

Downloads · 30 days

0

0% of all-time downloads

All-time downloads

24

Public

Repo size

279 KB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.mp4257 KB · 45%

At a glance

Task
Reinforcement Learning
Library
stable-baselines3
Access
Public
Created
Dec 11, 2022
Updated
Dec 11, 2022
SHA
c7a37c6f
Task
Reinforcement Learning
Library
stable-baselines3
Created
Dec 11, 2022
Updated
Dec 11, 2022