Downloads · 30 days
42
70% of all-time downloads
Guru-Raja-124/ppo-LunarLander-v2
ppo-LunarLander-v2 is a reinforcement learning model from Guru-Raja-124. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. It is set up for stable-baselines3.
Stable-Baselines3 PPO agent trained on LunarLander-v2.
Downloads · 30 days
42
70% of all-time downloads
All-time downloads
60
Public
Repo size
680 KB
Likes
0
Public
Click a slice to open those files.
.zip154 KB · 36%
From the Hugging Face model README
Stable-Baselines3 PPO agent trained on LunarLander-v2.
mean_reward - std_reward): 233.25This model was trained using Stable-Baselines3 PPO with an MLP policy.