Skip to content

Dash10107

LunarLander-v2

Dash10107/LunarLander-v2

LunarLander-v2 is a reinforcement learning model from Dash10107. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product.

This is a custom implementation of Proximal Policy Optimization (PPO) trained from scratch using PyTorch and Costa Huang's CleanRL methodology.

Downloads · 30 days

6

20% of all-time downloads

All-time downloads

30

Public

Repo size

936 KB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.mp4195 KB · 37%

At a glance

Task
Reinforcement Learning
Access
Public
Created
Apr 12, 2026
Updated
Apr 12, 2026
SHA
2fe75f3a
Task
Reinforcement Learning
Created
Apr 12, 2026
Updated
Apr 12, 2026