Downloads · 30 days
11
23% of all-time downloads
interview-eval/zephyr-7b-math-train-test
zephyr-7b-math-train-test is a text generation model from interview-eval. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
11
23% of all-time downloads
All-time downloads
47
Public
Parameters
7.2B
14.5 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors14.5 GB · 100%
From the Hugging Face model README
This model is a fine-tuned version of alignment-handbook/zephyr-7b-sft-full on the EunsuKim/MATH dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 0.746 | 1.0 | 10 | 0.6231 |
| 0.5258 | 2.0 | 20 | 0.3843 |
| 0.3183 | 3.0 | 30 | 0.1999 |
| 0.16 | 4.0 | 40 | 0.0864 |
| 0.0811 | 5.0 | 50 | 0.0496 |
| 0.0502 | 6.0 | 60 | 0.0345 |
| 0.035 | 7.0 | 70 | 0.0254 |
| 0.0241 | 8.0 | 80 | 0.0185 |
| 0.0165 | 9.0 | 90 | 0.0142 |
| 0.0129 | 10.0 | 100 | 0.0130 |