Downloads · 30 days
1
1% of all-time downloads
DevilNReality/ppo-LunarLander-v3
ppo-LunarLander-v3 is a reinforcement learning model from DevilNReality. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. It is set up for stable-baselines3.
This project trains a reinforcement learning agent using Proximal Policy Optimization (PPO) to solve the LunarLander-v3 environment from Gymnasium. The agent learns to land a lunar module safely between two flags usin…
Downloads · 30 days
1
1% of all-time downloads
All-time downloads
176
Public
Repo size
1.2 MB
Likes
0
Public
Click a slice to open those files.
.zip150 KB · 41%
From the Hugging Face model README
This project trains a reinforcement learning agent using Proximal Policy Optimization (PPO) to solve the LunarLander-v3 environment from Gymnasium. The agent learns to land a lunar module safely between two flags using Box2D physics.
The trained model is available at:
🔗 DevilNReality/ppo-LunarLander-v3
Includes:
Watch the agent land in the environment:
This model satisfies the Hugging Face Deep RL course Unit 1 requirements:
LunarLander-v3