Skip to content

Jereeli

ppo-LunarLander-v3

Jereeli/ppo-LunarLander-v3

ppo-LunarLander-v3 is a reinforcement learning model from Jereeli. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. It is set up for stable-baselines3.

This is a trained model of a PPO agent playing LunarLander-v3 using the stable-baselines3 library. Trained for ~1.1M timesteps in two stages (initial training + continued fine-tuning). Mean evaluation reward over 10 d…

Downloads · 30 days

0

0% of all-time downloads

All-time downloads

37

Public

Repo size

728 KB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.mp4163 KB · 34%

At a glance

Task
Reinforcement Learning
Library
stable-baselines3
Access
Public
Created
Aug 28, 2026
Updated
Aug 30, 2026
SHA
f65d40a6
Task
Reinforcement Learning
Library
stable-baselines3
Created
Aug 28, 2026
Updated
Aug 30, 2026
ppo-LunarLander-v3 — AI Model — AIMarketly