Downloads · 30 days
46
1% of all-time downloads
qingy2024/Qwen2.5-Math-14B-Instruct-Preview
Qwen2.5-Math-14B-Instruct-Preview is a text generation model from qingy2024. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
- Developed by: qingy2019 - License: apache-2.0 - Finetuned from model : unsloth/qwen2.5-14b-instruct-bnb-4bit
Downloads · 30 days
46
1% of all-time downloads
All-time downloads
3.5K
Public
Parameters
14.8B
29.6 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors29.5 GB · 100%
From the Hugging Face model README
This Qwen 2.5 model was trained 2x faster with Unsloth and Huggingface's TRL library.
I fine-tuned it for 400 steps on garage-bAInd/Open-Platypus with a batch size of 3.
Detailed results can be found here
| Metric | Value |
|---|---|
| Avg. | 36.71 |
| IFEval (0-Shot) | 60.66 |
| BBH (3-Shot) | 47.02 |
| MATH Lvl 5 (4-Shot) | 28.47 |
| GPQA (0-shot) | 16.33 |
| MuSR (0-shot) | 19.63 |
| MMLU-PRO (5-shot) | 48.12 |