Skip to content

herman66

Qwen2.5-0.5B-Open-R1-Distill

herman66/Qwen2.5-0.5B-Open-R1-Distill

Qwen2.5-0.5B-Open-R1-Distill is a text generation model from herman66. Use it when you need the model to write or continue text. It is set up for transformers.

This model is a fine-tuned version of Qwen/Qwen2.5-0.5B-Instruct on the HuggingFaceH4/Bespoke-Stratos-17k dataset. It has been trained using TRL.

Downloads · 30 days

25

9% of all-time downloads

All-time downloads

278

Public

Parameters

494M

1000 MB on disk

Likes

1

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors988 MB · 98%

At a glance

Task
Text Generation
Library
transformers
Model type
qwen2
Access
Public
Created
Feb 7, 2025
Updated
Apr 28, 2025
SHA
f15b3241

Base models

Task
Text Generation
Library
transformers
Type
qwen2
Languages
zho, eng, fra, spa, por, deu
Created
Feb 7, 2025
Updated
Apr 28, 2025