Downloads · 30 days
0
HugThang/llama3-8B-GRPO
llama3-8B-GRPO is a machine learning model from HugThang. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
Downloads · 30 days
0
Access
Public
Updated May 6, 2025
Repo size
524 MB
Likes
0
Public
Click a slice to open those files.
.safetensors671 MB · 77%