Downloads · 30 days
28
60% of all-time downloads
Vibudhbh/lander-ppo_rl
lander-ppo_rl is a reinforcement learning model from Vibudhbh. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. It is set up for stable-baselines3.
This is a trained PPO agent that learned to land a spacecraft on the moon!
Downloads · 30 days
28
60% of all-time downloads
All-time downloads
47
Public
Repo size
149 KB
Likes
0
Public
Click a slice to open those files.
.zip149 KB · 98%
From the Hugging Face model README
This is a trained PPO agent that learned to land a spacecraft on the moon!
from stable_baselines3 import PPO
import gymnasium as gym
# Load the trained model
model = PPO.load("lunar_lander_ppo_model")
# Create environment
env = gym.make('LunarLander-v2', render_mode='human')
# Test the agent
obs, _ = env.reset()
for _ in range(1000):
action, _ = model.predict(obs, deterministic=True)
obs, reward, terminated, truncated, info = env.step(action)
if terminated or truncated:
obs, _ = env.reset()
env.close()
The agent was trained using PPO with the following hyperparameters:
The agent successfully learned to:
Watch it land on the moon! 🌙