Downloads · 30 days
19
22% of all-time downloads
AGI4Good/HSPMATH-7B
HSPMATH-7B is a text generation model from AGI4Good. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as llama2.
We have released HSPMATH-7B, a supervised fine-tuning model for MATH.
Downloads · 30 days
19
22% of all-time downloads
All-time downloads
85
Public
Parameters
6.7B
13.5 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors13.5 GB · 100%
From the Hugging Face model README
We have released HSPMATH-7B, a supervised fine-tuning model for MATH.
We constructed a supervised fine-tuning dataset of 75k samples through a simple yet effective method based on the MetaMathQA dataset. After supervised fine-tuning the Llemma-7B model, we achieved a strong performance of 64.3% on the GSM8K dataset. The dataset construction method involves introducing a hint before the solution. For details, refer to the paper: Hint-before-Solving Prompting: Guiding LLMs to Effectively Utilize Encoded Knowledge.
A comparison of performances with methods of similar model sizes (7B) is shown in the table below:
| Open-source Model (7B) | GSM8k |
|---|---|
| MetaMath-Mistral-7B | 77.7 |
| MetaMath-7B-V1.0 | 66.5 |
| HSPMATH-7B | 64.3 |
| Llemma-7B (SFT) | 58.7 |
| WizardMath-7B | 54.9 |
| RFT-7B | 50.3 |
| Qwen-7b | 47.84 |
| Mistral-7b | 37.83 |
| Yi-6b | 32.6 |
| ChatGLM-6B | 32.4 |
| LLaMA2-7b | 12.96 |
| Close-source Model | GSM8k |
|---|---|
| GPT-3.5 | 57.1 |
| PaLM-540B | 56.5 |
| Minerva-540B | 58.8 |
| Minerva-62B | 52.4 |
| Chinchilla-70B | 43.7 |
Note: