Downloads · 30 days
0
amanm10000/sprout
sprout is a text generation model from amanm10000. Use it when you need the model to write or continue text.
A small English story-completion transformer trained from random initialization on a prefix of TinyStories. Not a chatbot or instruction-following model.
Downloads · 30 days
0
Access
Public
Updated Sep 17, 2026
Repo size
130 MB
Likes
0
Public
Click a slice to open those files.
.pt130 MB · 100%
From the Hugging Face model README
A small English story-completion transformer trained from random initialization on a prefix of TinyStories. Not a chatbot or instruction-following model.
Code and training artifacts: https://github.com/amanmprojects/sprout
Validation uses 24,576 tokens in fixed sampled windows from the official validation split, not the entire validation corpus. The same sample selected the best checkpoint. No independent test benchmark was run. See FINDINGS.md for the recipe, samples, and limitations.
python -m pip install torch==2.8.0 tokenizers==0.23.2 huggingface_hub
python -c "from huggingface_hub import snapshot_download; snapshot_download('amanm10000/sprout', local_dir='sprout')"
cd sprout
python sample.py --checkpoint best.pt --tokenizer tokenizer.json --prompt "Once upon a time, there was a little fox who lived in the forest."
CPU is the default. Add --device cuda for a bf16-capable CUDA GPU. This checkpoint uses the included custom PyTorch implementation; it is not directly loadable via Transformers AutoModel or a hosted Inference API. Review downloaded Python code before executing it.
best.pt contains float32 model weights, architecture configuration, tokenizer hash, and checkpoint metadata. Optimizer state was removed; this release is for inference, not exact training resume.
Learns simple grammar and story structure, but invents props, changes pronouns and speaker roles, and loses plot continuity. May generate biased, unsafe, or incorrect text. It has not been safety-aligned and should not be used as an unsupervised children's application or a factual assistant.
Source dataset: TinyStories, with upstream CDLA-Sharing-1.0 terms. Raw training data is not redistributed here. No separate blanket license grant is asserted for the model in this release; review upstream terms and contact the repository owner for licensing questions.