Skip to content

RationalPursuit

Qwen3-4B-R1-SFT

RationalPursuit/Qwen3-4B-R1-SFT

Qwen3-4B-R1-SFT is a text generation model from RationalPursuit. Use it when you need the model to write or continue text. It is set up for peft.

This is a a experimental research artifact only. Trained on rasbt/mathdistill/data/deepseek-r1-math-train4000.json

Downloads · 30 days

0

0% of all-time downloads

All-time downloads

21

Public

Parameters

4B

12.3 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors8 GB · 100%

At a glance

Task
Text Generation
Library
peft
Model type
qwen3
Access
Public
Created
Jul 13, 2026
Updated
Jul 13, 2026
SHA
5706d347

Try a prompt

Base models

Task
Text Generation
Library
peft
Type
qwen3
Created
Jul 13, 2026
Updated
Jul 13, 2026