Skip to content

herry12sf

cliffwalking-qlearning

herry12sf/cliffwalking-qlearning

cliffwalking-qlearning is a reinforcement learning model from herry12sf. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product.

Trained on CliffWalking-v1 using tabular Q-Learning from scratch. - Observation space: Discrete(48) - Action space: Discrete(4) - Mean Reward: -13.00 ± 0.00 - Episodes: 100,000 - Learning rate: 0.7 | Gamma: 0.95 | Eps…

Downloads · 30 days

0

Access

Public

Updated Jun 10, 2026

Repo size

1.7 KB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.png76 KB · 51%

At a glance

Task
Reinforcement Learning
Access
Public
Created
Jun 10, 2026
Updated
Jun 10, 2026
SHA
5245e53e
Task
Reinforcement Learning
Created
Jun 10, 2026
Updated
Jun 10, 2026
cliffwalking-qlearning — AI Model — AIMarketly