Downloads · 30 days
0
HallD/SkeptiSTEM-4B-stageR1-lora
SkeptiSTEM-4B-stageR1-lora is a machine learning model from HallD. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
This repo contains the LoRA adapter for SkeptiSTEM-4B, fine-tuned from unsloth/Qwen3-4B-Base.
Downloads · 30 days
0
Access
Public
Updated Dec 17, 2025
Repo size
276 MB
Likes
0
Public
Click a slice to open those files.
.safetensors264 MB · 94%
From the Hugging Face model README
This repo contains the LoRA adapter for SkeptiSTEM-4B, fine-tuned from unsloth/Qwen3-4B-Base.
Stage: R1 STEM SFT (math + science + coding mixture).