Downloads · 30 days
101
13% of all-time downloads
theprint/PyRe-3B-v2-GGUF
PyRe-3B-v2-GGUF is a machine learning model from theprint. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.
Please note that this model is a WIP experiment into GRPO fine tuning on Python code problems for reasoning. The performance of this model varies greatly depending on task, prompt and parameters.
Downloads · 30 days
101
13% of all-time downloads
All-time downloads
793
Public
Repo size
23.8 GB
Likes
0
Public
Click a slice to open those files.
.gguf23.8 GB · 100%
From the Hugging Face model README
Please note that this model is a WIP experiment into GRPO fine tuning on Python code problems for reasoning. The performance of this model varies greatly depending on task, prompt and parameters.
I recommend a very low temperature, like 0.1. You may also see more consistent results by encouraging the use of <think> and <answer> tags in the system prompt.
Think through complex problems carefully, before giving the user your final answer. Use <think> and </think> to encapsulate your thoughts.
This GGUF is based on theprint/PyRe-3B-v2.
This llama model was trained 2x faster with Unsloth and Huggingface's TRL library.