Downloads · 30 days
0
ccmusic-database/chest_falsetto
chest_falsetto is a audio classification model from ccmusic-database. Use it for the audio classification task on the model card, and read the license before you ship it in a product. The card lists the license as mit.
The design of the chest-falsetto voice discrimination model aims to effectively differentiate between real and synthetic voices in audio samples, with four specific categories including male chest, male falsetto, fema…
Downloads · 30 days
0
Access
Public
Updated Feb 27, 2026
Repo size
2.3 GB
Likes
26
Public
Click a slice to open those files.
.pt515 MB · 100%
From the Hugging Face model README
The design of the chest-falsetto voice discrimination model aims to effectively differentiate between real and synthetic voices in audio samples, with four specific categories including male chest, male falsetto, female chest, and female falsetto voices. The model's training is based on a backbone network from the computer vision (CV) domain, which involves transforming audio data into spectrograms and fine-tuning to enhance the network's accuracy in recognizing different voice categories. During training, a dataset containing both real and synthetic voice samples is utilized to ensure the model adequately learns and captures features relevant to male and female chest and falsetto voices. Through this approach, the model can finely classify different genders and chest/falsetto voices, providing a reliable solution for accurate voice discrimination in audio. This model holds broad potential applications in fields such as speech processing and music production, offering an efficient and precise tool for audio analysis and processing. Its training and fine-tuning strategies based on computer vision principles highlight the model's adaptability and robustness across different domains, providing beneficial examples for further research and application.
https://huggingface.co/spaces/ccmusic-database/chest_falsetto
from huggingface_hub import snapshot_download
model_dir = snapshot_download(
"ccmusic-database/chest_falsetto",
cache_dir="./__pycache__",
)
print(model_dir)
GIT_LFS_SKIP_SMUDGE=1 git clone [email protected]:ccmusic-database/chest_falsetto
cd chest_falsetto
| Backbone | Mel | CQT | Chroma |
|---|---|---|---|
| Swin-S V2 | 0.968 | 0.268 | 0.268 |
| MaxViT-T | 0.820 | 0.933 | 0.250 |
| AlexNet | 0.994 | 0.963 | 0.586 |
| ShuffleNet V2 2.0 | 0.939 | 0.669 | 0.222 |
| GoogleNet | 0.983 | 0.274 | 0.292 |
| MNASNet-A3 | 0.756 | 0.260 | 0.320 |
| SqueezeNet 1.1 | 0.963 | 0.900 | 0.378 |
| Average | 0.918 | 0.610 | 0.331 |
https://huggingface.co/datasets/ccmusic-database/chest_falsetto
https://www.modelscope.cn/models/ccmusic-database/chest_falsetto
https://github.com/monetjoe/ccmusic_eval
@dataset{zhaorui_liu_2021_5676893,
author = {Zhaorui Liu and Zijin Li},
title = {Music Data Sharing Platform for Computational Musicology Research (CCMUSIC DATASET)},
month = nov,
year = 2021,
publisher = {Zenodo},
version = {1.1},
doi = {10.5281/zenodo.5676893},
url = {https://doi.org/10.5281/zenodo.5676893}
}