Downloads · 30 days
0
S-teven/tinygoop-1.1b
tinygoop-1.1b is a question answering model from S-teven. Use it when the input is a question plus a passage. The card lists the license as mit.
A fine-tuned version of TinyLlama-1.1B-Chat with room temp iq - quantized to 4 bits and trained on copypastas
Downloads · 30 days
0
Access
Public
Updated Nov 21, 2025
Repo size
2.2 GB
Likes
0
Public
Click a slice to open those files.
.safetensors2.2 GB · 100%
From the Hugging Face model README
A fine-tuned version of TinyLlama-1.1B-Chat with room temp iq -> quantized to 4 bits and trained on copypastas
Sources:
import torch
from transformers import AutoTokenizer, AutoModelForCausalLM
model_id = "S-teven/tinygoop-1.1b"
tokenizer = AutoTokenizer.from_pretrained(model_id)
model = AutoModelForCausalLM.from_pretrained(
model_id,
torch_dtype=torch.float16,
device_map="auto"
)
prompt = "hey"
inputs = tokenizer(prompt, return_tensors="pt").to(model.device)
outputs = model.generate(
**inputs,
max_new_tokens=256,
do_sample=True,
temperature=1.2,
top_p=0.95,
repetition_penalty=1.05
)
print(tokenizer.decode(outputs[0], skip_special_tokens=True))
| Precision | VRAM Required | Hardware |
|---|---|---|
| 4-bit Quantized | ~800MB | Any modern GPU |
| CPU (FP32) | ~4GB RAM | Modern CPU (slow) |
Content Warning: This model was trained on copypasta data and may generate:
Not suitable for: