Downloads · 30 days
0
SamTru88/ppo-CleanRL-LunarLander
ppo-CleanRL-LunarLander is a reinforcement learning model from SamTru88. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product.
Downloads · 30 days
0
Access
Public
Updated Jul 24, 2026
Repo size
43.3 KB
Likes
0
Public
Click a slice to open those files.
.pt43.3 KB · 95%
From the Hugging Face model README
Mean reward: 51.05 +/- 98.11