Downloads · 30 days
23
11% of all-time downloads
EpistemeAI/metatune-gpt20b-R1
metatune-gpt20b-R1 is a text generation model from EpistemeAI. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
- Generates new data for itself, - Evaluates its performance, and - Adjusts its own hyperparameters based on improvement metrics.
Downloads · 30 days
23
11% of all-time downloads
All-time downloads
213
Public
Parameters
21.5B
13.8 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors13.8 GB · 100%
How the weights are stored.
U819.7B · 92%
From the Hugging Face model README
Due to recursive self improvement method, there is no final model, but improved model, this is a 5th metacycle(generation) improved checkpoint model.
You can use gpt-oss-120b and gpt-oss-20b with Transformers. If you use the Transformers chat template, it will automatically apply the harmony response format. If you use model.generate directly, you need to apply the harmony format manually using the chat template or use our openai-harmony package.
To get started, install the necessary dependencies to setup your environment:
pip install -U transformers kernels torch
For Google Colab (free/Pro)
!pip install -q --upgrade torch
!pip install -q transformers triton==3.4 kernels
!pip uninstall -q torchvision torchaudio -y
Once, setup you can proceed to run the model by running the snippet below:
from transformers import pipeline
import torch
model_id = "EpistemeAI/metatune-gpt20b-R1"
pipe = pipeline(
"text-generation",
model=model_id,
torch_dtype="auto",
device_map="auto",
)
messages = [
{"role": "user", "content": "Derive the Euler–Lagrange equation from the principle of stationary action.""},
]
outputs = pipe(
messages,
max_new_tokens=3000,
)
print(outputs[0]["generated_text"][-1])
You can adjust the reasoning level that suits your task across three levels:
The reasoning level can be set in the system prompts, e.g., "Reasoning: high".
The gpt-oss models are excellent for:
Both gpt-oss models can be fine-tuned for a variety of specialized use cases.
This smaller model gpt-oss-20b can be fine-tuned on consumer hardware, whereas the larger gpt-oss-120b can be fine-tuned on a single H100 node.
| Tasks | metatune R0 | metatune R1 | Llama 4 Maverick |
|---|---|---|---|
| gsm8k_cot | 0.91 | 0.9796 | - |
| gpqa_diamond_cot_n_shot | 0.722 | - | |
| hellaswag | 0.421 | 0.525 | - |
| arc_challenge | 0.349 | 0.349 | - |
| winogrande | 0.7851 | 0.5928 | - |
Jurgen Schmidhuber
This gpt_oss model was trained 2x faster with Unsloth and Huggingface's TRL library.