Downloads · 30 days
18
9% of all-time downloads
evilyesh/Qwen2.5-Coder-14B-Instruct-Thinking
Qwen2.5-Coder-14B-Instruct-Thinking is a machine learning model from evilyesh. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
This model is a fine-tuned version of the base model Qwen/Qwen2.5-Coder-14B-Instruct. It was trained on a subset of problems from the GAIR/LIMO dataset, specifically focusing on 611 problems over 14 training epochs.
Downloads · 30 days
18
9% of all-time downloads
All-time downloads
205
Public
Parameters
14.8B
354 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors29.5 GB · 100%
From the Hugging Face model README
This model is a fine-tuned version of the base model Qwen/Qwen2.5-Coder-14B-Instruct. It was trained on a subset of problems from the GAIR/LIMO dataset, specifically focusing on 611 problems over 14 training epochs.
After testing more I found that the model does not always include reasoning, I will update with more epochs.
Warning! The model often goes into an endless chain of reasoning. Need to use phrase: "Structure your thoughts. Be attentive to details."
During testing, the fine-tuned model demonstrated significant improvements in reasoning ability compared to the base model. It began to provide more coherent and accurate responses, avoiding the mistakes observed in the base model during my initial tests.
While preliminary results are promising, further evaluation is needed to assess the overall improvement in model quality. I encourage the community to test the model and share their findings. Your feedback will be invaluable in understanding the extent of the improvements.
Special thanks to the creators of the Qwen/Qwen2.5-Coder-14B-Instruct and GAIR/LIMO datasets for providing the foundational resources.