Downloads · 30 days
0
tanu20/ppo-lunarlander
ppo-lunarlander is a reinforcement learning model from tanu20. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product.
This repository contains a PPO (Proximal Policy Optimization) agent implemented in PyTorch and trained on the LunarLander environment as part of the Hugging Face Deep Reinforcement Learning Course — Unit 8 Part I.
Downloads · 30 days
0
Access
Public
Updated Sep 23, 2026
Repo size
321 KB
Likes
0
Public
Click a slice to open those files.
.mp4234 KB · 72%
From the Hugging Face model README
This repository contains a PPO (Proximal Policy Optimization) agent implemented in PyTorch and trained on the LunarLander environment as part of the Hugging Face Deep Reinforcement Learning Course — Unit 8 Part I.
LunarLander-v3Evaluation was performed over 10 episodes.
Mean reward: -132.37 +/- 43.43
model.pt — trained PPO modelresults.json — evaluation resultsreplay.mp4 — agent gameplay replayHugging Face Deep Reinforcement Learning Course
Unit 8 Part I — PPO with PyTorch.