Skip to content

imflash217

proximal_policy_optimization_lunar_lander_v2

imflash217/proximal_policy_optimization_lunar_lander_v2

proximal_policy_optimization_lunar_lander_v2 is a reinforcement learning model from imflash217. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. It is set up for stable-baselines3.

This is a trained model of a Proximal Parity Optimization (PPO) agent playing LunarLander-v2 using the stable-baselines3 library.

Downloads · 30 days

2

8% of all-time downloads

All-time downloads

26

Public

Repo size

279 KB

Likes

4

Trending 2

Hugging Face

Repo makeup

Click a slice to open those files.

.mp4215 KB · 41%

At a glance

Task
Reinforcement Learning
Library
stable-baselines3
Access
Public
Created
Jan 13, 2023
Updated
Jan 13, 2023
SHA
ddbf5906
Task
Reinforcement Learning
Library
stable-baselines3
Created
Jan 13, 2023
Updated
Jan 13, 2023