Downloads · 30 days
0
MattStammers/ppo-LunarLander-v2-fullcoded
ppo-LunarLander-v2-fullcoded is a reinforcement learning model from MattStammers. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product.
This is a trained model of a PPO agent playing LunarLander-v2.
Downloads · 30 days
0
Access
Public
Updated Sep 19, 2023
Repo size
10.2 MB
Likes
0
Public
Click a slice to open those files.
.09.4 MB · 97%
From the Hugging Face model README
This is a trained model of a PPO agent playing LunarLander-v2.
See the GitHub for full info and the journey on creating this on the surface not particularly exciting model: https://github.com/MattStammers/PPO_Lander_Implementation
It took me 8 attempts to get the score to nearly reach 0 using a cleanRL implementation and WandB metric tracking and then this version was trained after 10 attempts converging at about 3 million training steps