Downloads · 30 days
35
100% of all-time downloads
0GiS0/lunarlander-v3
lunarlander-v3 is a reinforcement learning model from 0GiS0. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. It is set up for stable-baselines3.
Entrenado con Stable-Baselines3 PPO y una política MLP.
Downloads · 30 days
35
100% of all-time downloads
All-time downloads
35
Public
Repo size
300 KB
Likes
0
Public
Click a slice to open those files.
.zip150 KB · 67%
From the Hugging Face model README
Entrenado con Stable-Baselines3 PPO y una política MLP.
Recompensa media en 10 episodios deterministas: 252.88 +/- 17.76.
<video controls src="replay.mp4"></video>
Carga solo checkpoints de fuentes en las que confíes.
from huggingface_hub import hf_hub_download
from stable_baselines3 import PPO
checkpoint = hf_hub_download(repo_id='0GiS0/lunarlander-v3', filename='ppo-LunarLander-v3.zip')
model = PPO.load(checkpoint, device='cpu')