Downloads · 30 days
87
9% of all-time downloads
tencent/SongPrep-7B
SongPrep-7B is a automatic speech recognition model from tencent. Use it when you need speech turned into text.
<p align="center"<img src="img/logo.jpg" width="40%"</p <p align="center" <a href="https://song-prep.github.io/demo/"Demo</a | <a href="https://arxiv.org/abs/2509.17404"Paper</a | <a href="http…
Downloads · 30 days
87
9% of all-time downloads
All-time downloads
958
Public
Parameters
7.7B
24.8 GB on disk
Likes
46
Public
Click a slice to open those files.
.safetensors16.1 GB · 100%
From the Hugging Face model README
This repository is the official weight repository for SongPrep: A Preprocessing Framework and End-to-end Model for Full-song Structure Parsing and Lyrics Transcription. In this repository, we provide the SongPrep-7B model that has been trained on the Million Song Dataset.
| Model | #Params | HuggingFace |
|---|---|---|
| SongPrep | 7B | you are here |
@misc{tan2025songpreppreprocessingframeworkendtoend,
title={SongPrep: A Preprocessing Framework and End-to-end Model for Full-song Structure Parsing and Lyrics Transcription},
author={Wei Tan and Shun Lei and Huaicheng Zhang and Guangzheng Li and Yixuan Zhang and Hangting Chen and Jianwei Yu and Rongzhi Gu and Dong Yu},
year={2025},
eprint={2509.17404},
archivePrefix={arXiv},
primaryClass={eess.AS},
url={https://arxiv.org/abs/2509.17404},
}
The code and weights in this repository is released in the LICENSE file.