Downloads · 30 days
0
prashanthbsp/grpo_saved_lora
grpo_saved_lora is a machine learning model from prashanthbsp. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.
- Developed by: prashanthbsp - License: apache-2.0 - Finetuned from model : unsloth/Qwen3-4B-Base
Downloads · 30 days
0
Access
Public
Updated May 29, 2025
Repo size
264 MB
Likes
0
Public
Click a slice to open those files.
.safetensors264 MB · 100%
From the Hugging Face model README
This qwen3 model was trained 2x faster with Unsloth and Huggingface's TRL library.