Downloads · 30 days
0
herry12sf/cliffwalking-qlearning
cliffwalking-qlearning is a reinforcement learning model from herry12sf. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product.
Trained on CliffWalking-v1 using tabular Q-Learning from scratch. - Observation space: Discrete(48) - Action space: Discrete(4) - Mean Reward: -13.00 ± 0.00 - Episodes: 100,000 - Learning rate: 0.7 | Gamma: 0.95 | Eps…
Downloads · 30 days
0
Access
Public
Updated Jun 10, 2026
Repo size
1.7 KB
Likes
0
Public
Click a slice to open those files.
.png76 KB · 51%
From the Hugging Face model README
Trained on CliffWalking-v1 using tabular Q-Learning from scratch.