Downloads · 30 days
62
14% of all-time downloads
artindnr/ChatBerry-1
ChatBerry-1 is a text generation model from artindnr. Use it when you need the model to write or continue text. The card lists the license as apache-2.0.
Downloads · 30 days
62
14% of all-time downloads
All-time downloads
442
Public
Parameters
20.9B
41.9 GB on disk
Likes
2
Public
Click a slice to open those files.
.safetensors41.8 GB · 100%
How the weights are stored.
F1620.9B · 100%
From the Hugging Face model README

ChatBerry-1 is a fine-tuned version of artindnr/strawberry-1, converted from a reasoning ("thinking") model into a direct-answer chat model. Reasoning traces are disabled — ChatBerry-1 responds directly, without emitting a separate chain-of-thought / analysis channel, making it faster and simpler to deploy for everyday conversational use in Farsi, English, and other languages.
gpt_ossStrawberry-1 was built to generate high-quality Farsi reasoning traces. ChatBerry-1 takes the opposite approach: it strips reasoning out of the loop entirely, so the model:
analysis (chain-of-thought) channelChatBerry-1 uses the gpt-oss chat template (Harmony format) shipped with the base model, so it works with 🤗 Transformers.
pip install torch --index-url https://download.pytorch.org/whl/cu128
pip install "trl>=0.20.0" "peft>=0.17.0" "transformers>=4.55.0" "kernels>=0.12.0"
This has been verified to work with:
| Package | Version |
|---|---|
torch | 2.8.0+cu129 |
transformers | 5.14.1 |
trl | 1.9.2 |
peft | 0.20.0 |
accelerate | 1.10.1 |
tokenizers | 0.22.0 |
import torch
from transformers import AutoModelForCausalLM, AutoTokenizer
MODEL_ID = "artindnr/chatberry-1"
tokenizer = AutoTokenizer.from_pretrained(MODEL_ID)
model = AutoModelForCausalLM.from_pretrained(
MODEL_ID,
torch_dtype=torch.bfloat16,
device_map="auto",
)
USER_PROMPT = "تو کی هستی و اسمت چیه؟"
messages = [
{"role": "user", "content": USER_PROMPT},
]
inputs = tokenizer.apply_chat_template(
messages,
add_generation_prompt=True,
tokenize=True,
return_dict=True,
return_tensors="pt",
).to(model.device)
outputs = model.generate(
**inputs,
max_new_tokens=512,
temperature=0.6,
do_sample=True,
)
print(tokenizer.decode(outputs[0][inputs["input_ids"].shape[-1]:], skip_special_tokens=True))
Unlike Strawberry-1, there's no need to set a reasoning language system message or parse out separate analysis / final channels — ChatBerry-1 goes straight to its final answer in the final channel, so decoding just the newly generated tokens with skip_special_tokens=True gives you the plain-text response directly.
ChatBerry-1 is intended for:
artindnr/strawberry-1 may be a better fit.gpt-oss-20b base model and the strawberry-1 checkpoint it was built from, including the possibility of hallucinated facts.This model is released under the Apache 2.0 license, consistent with the base gpt-oss-20b model and strawberry-1.
If you use ChatBerry-1 in your work, please cite:
@misc{chatberry1,
title = {ChatBerry-1: A Direct-Answer Chat Fine-tune of Strawberry-1},
author = {artindnr},
year = {2026},
url = {https://huggingface.co/artindnr/chatberry-1}
}
Built on top of artindnr/strawberry-1, itself fine-tuned from openai/gpt-oss-20b