Downloads · 30 days
0
0% of all-time downloads
RationalPursuit/Qwen3-4B-R1-SFT
Qwen3-4B-R1-SFT is a text generation model from RationalPursuit. Use it when you need the model to write or continue text. It is set up for peft.
This is a a experimental research artifact only. Trained on rasbt/mathdistill/data/deepseek-r1-math-train4000.json
Downloads · 30 days
0
0% of all-time downloads
All-time downloads
21
Public
Parameters
4B
12.3 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors8 GB · 100%
From the Hugging Face model README
This is a a experimental research artifact only. Trained on rasbt/math_distill/data/deepseek-r1-math-train_4000.json
THIS IS RESEARCH ARTIFACT and should not intended for use.