Downloads · 30 days
28
50% of all-time downloads
koutch/paper_qwen_2.json_train_grpo_v1_train_code
paper_qwen_2.json_train_grpo_v1_train_code is a text generation model from koutch. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
- Developed by: koutch - License: apache-2.0 - Finetuned from model : unsloth/qwen3-4b-instruct-2507-unsloth-bnb-4bit
Downloads · 30 days
28
50% of all-time downloads
All-time downloads
56
Public
Parameters
4B
8.1 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors8 GB · 100%
From the Hugging Face model README
This qwen3 model was trained 2x faster with Unsloth and Huggingface's TRL library.