Downloads · 30 days
0
JackYoung27/writesae-ckpts
writesae-ckpts is a machine learning model from JackYoung27. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for pytorch. The card lists the license as mit.
WriteSAE uses a sparse autoencoder to study what a recurrent language model stores in memory and how it affects predictions. This recurrent state is a matrix that carries information between tokens. We reconstruct sav…
Downloads · 30 days
0
Access
Public
Updated Sep 8, 2026
Repo size
538 MB
Likes
0
Public
Click a slice to open those files.
.pt511 MB · 87%
From the Hugging Face model README
WriteSAE uses a sparse autoencoder to study what a recurrent language model stores in memory and how it affects predictions. This recurrent state is a matrix that carries information between tokens. We reconstruct saved states from a few active features, each shaped like a native memory write.
python -m pip install huggingface_hub torch
hf download JackYoung27/writesae-ckpts --local-dir writesae --include 'core/*' 'LOAD_EXAMPLE.py' 'writesae/qwen0p8b/L9_H4/*'
cd writesae
python LOAD_EXAMPLE.py
This example loads the Qwen3.5-0.8B layer 9, head 4 autoencoder on a CPU and checks the output shapes.
writesae/: seven trained autoencoder packs, each with weights and settings.flat_baseline/: comparison checkpoints.core/: autoencoder and training code.experiments/ and scripts/: code for testing memory edits.results/: saved experiment outputs.manifest.json: artifact metadata and checksums.The loading example checks reconstruction. It does not run a memory-editing experiment or load the base language model.