Downloads · 30 days
77
28% of all-time downloads
theworker02/open-reason-large
open-reason-large is a text generation model from theworker02. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
This is a large GPT-2-style causal LM trained from scratch on the Open Reason SFT split. It is larger than theworker02/open-reason-medium (13,867,008 parameters) and is not a 1B model and is not theworker02/open-reaso…
Downloads · 30 days
77
28% of all-time downloads
All-time downloads
272
Public
Parameters
91.5M
366 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors366 MB · 100%
From the Hugging Face model README
This is a large GPT-2-style causal LM trained from scratch on the Open Reason SFT split. It is larger than theworker02/open-reason-medium (13,867,008 parameters) and is not a 1B model and is not theworker02/open-reason-1b.
Weights live on this Hub repo. They are not stored in the GitHub git tree.
n_layer=12, n_embd=768, n_head=12, vocab_size=8192, max_seq_len=256data/release/all.jsonltheworker02/open-reason pipeline 1.4.0torch 2.12.0+cpu; torch.cuda.is_available()=False; Docker not installed and not used; nvidia-smi not present. NVIDIA CUDA was not used. AMD GPU / ROCm / DirectML were not used.No Reddit sources.
from transformers import AutoModelForCausalLM, AutoTokenizer
tok = AutoTokenizer.from_pretrained("theworker02/open-reason-large")
model = AutoModelForCausalLM.from_pretrained("theworker02/open-reason-large")