Downloads · 30 days
0
Parth673/LunarLander-v2
LunarLander-v2 is a reinforcement learning model from Parth673. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product.
This is a trained model of a PPO agent playing LunarLander-v2.
Downloads · 30 days
0
Access
Public
Updated Dec 16, 2023
Repo size
2.6 MB
Likes
0
Public
Click a slice to open those files.
.02.4 MB · 91%
From the Hugging Face model README
This is a trained model of a PPO agent playing LunarLander-v2.
See the GitHub for full info and the journey on creating this on the surface not particularly exciting model: https://github.com/MattStammers/PPO_Lander_Implementation
It took me 8 attempts to get the score to nearly reach 0 using a cleanRL implementation and WandB metric tracking and then this version was trained after 10 attempts converging at about 3 million training steps