Downloads · 30 days
7
4% of all-time downloads
Jashan887/83_Self_Improving_Loop
83_Self_Improving_Loop is a text generation model from Jashan887. Use it when you need the model to write or continue text. It is set up for llama.cpp. The card lists the license as apache-2.0.
Stage 2 checkpoint for the Hermes/Bonsai Karpathy auto-research loop.
Downloads · 30 days
7
4% of all-time downloads
All-time downloads
170
Public
Repo size
—
Likes
0
Public
Click a slice to open those files.
.gguf8.7 GB · 100%
From the Hugging Face model README
Stage 2 checkpoint for the Hermes/Bonsai Karpathy auto-research loop.
Last updated: 2026-04-05
This release is inspired by Andrej Karpathy's framing of self-improving training loops and auto-research. It contains the model artifact that worked, plus a concise model card explaining how it was produced and how to run it.
Best performance concentrated in:
Trained via a graduation protocol with teacher-guided validation, raw-pass reinforcement, and frontier failure analysis. The interesting contribution is the loop methodology; see GitHub for the full curriculum and training workflow.
The training loop, curriculum design, graduation protocol, and detailed methodology live here:
https://github.com/aurous37-lang/Hermes-Bonsai-Self-Improving-Agent-Loop
bonsai-8b-stage2-post-curriculum-q8.gguf — the shipped stage 2 checkpointREADME.md — this model cardLICENSE — Apache-2.0 licenseRecommended working config from the stable local run:
--ctx-size 40960--n-gpu-layers 37./llama-cli -m bonsai-8b-stage2-post-curriculum-q8.gguf \
--ctx-size 40960 \
-p "Explain the CAP theorem for a backend engineer."
./llama-server -m bonsai-8b-stage2-post-curriculum-q8.gguf \
--ctx-size 40960 \
--n-gpu-layers 37 \
--host 0.0.0.0 --port 8080
Then point your client at the local OpenAI-compatible endpoint exposed by llama-server.