Downloads · 30 days
0
sotoalt/lepong
lepong is a reinforcement learning model from sotoalt. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. The card lists the license as mit.
A 13M-parameter JEPA world model that plays Pong by watching pixels.
Downloads · 30 days
0
Access
Public
Updated Apr 13, 2026
Repo size
164 MB
Likes
0
Public
Click a slice to open those files.
.pt158 MB · 96%
From the Hugging Face model README
A 13M-parameter JEPA world model that plays Pong by watching pixels.
The encoder + predictor are frozen (13M params). Only the state head trains (1,930 params).
| File | Description |
|---|---|
| lepong_statehead_occ_aug.pt | Shipping checkpoint - trained with occlusion augmentation |
| lepong_statehead_frozen.pt | Baseline - trained on unoccluded frames only |
| lepong_v1.pt | Init checkpoint - encoder + predictor only, no state head |
| pong_train_30k.npz | Training data - 30K frames (128x128 RGB) + states + actions |
| Metric | Value |
|---|---|
| ball_y median error (in-dist) | 2.8% |
| Controller success (in-dist) | 99.3% |
| Controller success (OOD) | 88.7% |
| ball_x improvement at 40% occ (augmented) | -58% |
Live demo: sotoalt.dev/experiments/lepong.html
Code: github.com/SotoAlt/lepong
MIT