Downloads · 30 days
19
25% of all-time downloads
iko-01/iko-002
iko-002 is a text generation model from iko-01. Use it when you need the model to write or continue text. The card lists the license as apache-2.0.
iko-2 is the second model in the iko series — a GPT-2 Medium (355M parameters) language model that combines:
Downloads · 30 days
19
25% of all-time downloads
All-time downloads
75
Public
Parameters
355M
710 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors710 MB · 100%
From the Hugging Face model README
iko-2 is the second model in the iko series — a GPT-2 Medium (355M parameters) language model that combines:
GPT-2 (124M) → [FineWeb fine-tune] → iko-1
↓ distillation
GPT-2 Medium (355M) → [QLoRA + Reddit + Replay] → [TIES merge] → iko-2
from transformers import AutoModelForCausalLM, AutoTokenizer
model = AutoModelForCausalLM.from_pretrained("iko-01/iko-002")
tokenizer = AutoTokenizer.from_pretrained("iko-01/iko-002")
input_text = "The best thing about learning is"
inputs = tokenizer(input_text, return_tensors="pt")
outputs = model.generate(**inputs, max_new_tokens=100, do_sample=True, temperature=0.8)
print(tokenizer.decode(outputs[0], skip_special_tokens=True))
| Model | Parameters | Training Data | Method |
|---|---|---|---|
| iko-1 | 124M | FineWeb (700K docs) | QLoRA on GPT-2 |
| iko-2 | 355M | Reddit + iko-1 distillation | QLoRA + TIES merge on GPT-2 Medium |
Apache 2.0