Skip to content

proyrb

ppo-LunarLander-v2

proyrb/ppo-LunarLander-v2

ppo-LunarLander-v2 is a reinforcement learning model from proyrb. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product.

This is a trained model of a PPO agent playing LunarLander-v2. - Mean Reward: -56.14 ± 76.83 - Number of Evaluation Episodes: 10 {'envid': 'LunarLander-v2' 'totaltimesteps': 100000 'learningrate': 0.0003 'numenvs': 8…

Downloads · 30 days

0

Access

Public

Updated Jun 16, 2025

Repo size

5.9 MB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.2f208e49a8655.6 MB · 99%

At a glance

Task
Reinforcement Learning
Access
Public
Created
Jun 16, 2025
Updated
Jun 16, 2025
SHA
7d77c5a5
Task
Reinforcement Learning
Created
Jun 16, 2025
Updated
Jun 16, 2025