Downloads · 30 days
0
mjob1/4b
4b is a reinforcement learning model from mjob1. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product.
This is a trained model of a Reinforce agent playing Pixelcopter-PLE-v0. To learn how to train your own models, check Unit 4 of the Deep Reinforcement Learning Course: Course Link
Downloads · 30 days
0
Access
Public
Updated Oct 25, 2024
Repo size
235 KB
Likes
0
Public
Click a slice to open those files.
.pth78.4 KB · 56%
From the Hugging Face model README
This is a trained model of a **Reinforce** agent playing **Pixelcopter-PLE-v0**.
To learn how to train your own models, check Unit 4 of the Deep Reinforcement Learning Course:
[Course Link](https://huggingface.co/deep-rl-course/unit4/introduction)