Downloads · 30 days
14
23% of all-time downloads
Alelcv27/Llama3.2-1B-Instruct-Math-GRPO
Llama3.2-1B-Instruct-Math-GRPO is a text generation model from Alelcv27. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
- Developed by: Alelcv27 - License: apache-2.0 - Finetuned from model : unsloth/Llama-3.2-1B
Downloads · 30 days
14
23% of all-time downloads
All-time downloads
60
Public
Parameters
1.2B
2.5 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors2.5 GB · 99%
From the Hugging Face model README
This llama model was trained 2x faster with Unsloth and Huggingface's TRL library.