Downloads · 30 days
9
26% of all-time downloads
lllqaq/R2EGym-7B-Agent-Coder-Instruct-filtered2
R2EGym-7B-Agent-Coder-Instruct-filtered2 is a machine learning model from lllqaq. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
This model is a full SFT fine-tune of Qwen/Qwen2.5-Coder-7B-Instruct on trajgpt5mini (filtered2).
Downloads · 30 days
9
26% of all-time downloads
All-time downloads
35
Public
Parameters
333K
15.2 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors15.2 GB · 100%
From the Hugging Face model README
This model is a full SFT fine-tune of Qwen/Qwen2.5-Coder-7B-Instruct on traj_gpt5mini (filtered2).
Export source checkpoint: /data/jiarong/LLaMA-Factory/saves/R2EGym-7B-Agent-Coder-Instruct2/checkpoint-188