Downloads · 30 days
50
17% of all-time downloads
Kuiyen/model-name_UnslothGRPO
model-name_UnslothGRPO is a machine learning model from Kuiyen. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
This model was finetuned and converted to GGUF format using Unsloth.
Downloads · 30 days
50
17% of all-time downloads
All-time downloads
286
Public
Parameters
1000M
4.8 GB on disk
Likes
0
Public
Click a slice to open those files.
.gguf2.7 GB · 57%
From the Hugging Face model README
This model was finetuned and converted to GGUF format using Unsloth.
Example usage:
llama-cli -hf Kuiyen/model-name_UnslothGRPO --jinjallama-mtmd-cli -hf Kuiyen/model-name_UnslothGRPO --jinjagemma-3-1b-it.Q5_K_M.ggufgemma-3-1b-it.Q8_0.ggufgemma-3-1b-it.Q4_K_M.ggufAn Ollama Modelfile is included for easy deployment.
The model's BOS token behavior was adjusted for GGUF compatibility. This was trained 2x faster with Unsloth <img src="https://raw.githubusercontent.com/unslothai/unsloth/main/images/unsloth%20made%20with%20love.png" width="200"/>