Downloads · 30 days
9
9% of all-time downloads
TOFU-SFT/Qwen2.5-Math-7B-4bit
Qwen2.5-Math-7B-4bit is a text generation model from TOFU-SFT. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
- Developed by: Qwen - Model type: Causal Language Models - License: Apache 2.0
Downloads · 30 days
9
9% of all-time downloads
All-time downloads
95
Public
Parameters
7.8B
5.6 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors5.5 GB · 100%
How the weights are stored.
U86.7B · 86%
From the Hugging Face model README
@article{yang2024qwen25mathtechnicalreportmathematical,
title={Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement},
author={An Yang and Beichen Zhang and Binyuan Hui and Bofei Gao and Bowen Yu and Chengpeng Li and Dayiheng Liu and Jianhong Tu and Jingren Zhou and Junyang Lin and Keming Lu and Mingfeng Xue and Runji Lin and Tianyu Liu and Xingzhang Ren and Zhenru Zhang},
journal={arXiv preprint arXiv:2409.12122},
year={2024}
}