Downloads · 30 days
424
16% of all-time downloads
littlelearner/littlelearner-5b-chatty
littlelearner-5b-chatty is a text generation model from littlelearner. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as other.
5B K-5-bounded chat model with general chat, model identity, and format steerability installed by a behavior SFT on the blend base (chatty v2).
Downloads · 30 days
424
16% of all-time downloads
All-time downloads
2.6K
Public
Parameters
5B
10.1 GB on disk
Likes
5
Public
Click a slice to open those files.
.safetensors10.1 GB · 100%
From the Hugging Face model README
5B K-5-bounded chat model with general chat, model identity, and format steerability installed by a behavior SFT on the blend base (chatty v2).
Part of the LittleLearner scale-up study (pedagogically-controlled knowledge exposure): Qwen3 dense LMs trained on a corpus filtered to U.S. K-5 material (bounded) vs an unfiltered FineWeb-Edu corpus (unbounded), to measure what an interpretable knowledge boundary costs and grants.
Project page: https://littlelearner-ll.github.io/
Qwen3ForCausalLM).MathCAMPS (paper-filtered):
from transformers import AutoModelForCausalLM, AutoTokenizer
repo = "littlelearner/littlelearner-5b-bounded-sft-chatty"
tok = AutoTokenizer.from_pretrained(repo)
model = AutoModelForCausalLM.from_pretrained(repo, dtype="bfloat16", device_map="auto")
msgs = [{"role": "user", "content": "If Sarah has 12 apples and gives 5 to Tom, how many does she have left?"}]
ids = tok.apply_chat_template(msgs, add_generation_prompt=True, return_tensors="pt").to(model.device)
out = model.generate(ids)
print(tok.decode(out[0, ids.shape[1]:], skip_special_tokens=True))
# vLLM
from vllm import LLM
repo = "littlelearner/littlelearner-5b-bounded-sft-chatty"
llm = LLM(repo)
msgs = [{"role": "user", "content": "Liam has 3 apples and buys 4 more. How many apples does he have?"}]
print(llm.chat(msgs)[0].outputs[0].text)