Downloads · 30 days
11
3% of all-time downloads
birgermoell/psst-fairseq-rir
psst-fairseq-rir is a automatic speech recognition model from birgermoell. Use it when you need speech turned into text. It is set up for transformers. The card lists the license as apache-2.0.
This model is trained on the PSST Challenge data, with a subset of TIMIT that was augmented using Room Impulse Response (RIR). A file containing the list of TIMIT IDs is in the repository (timit-ids.txt)
Downloads · 30 days
11
3% of all-time downloads
All-time downloads
391
Public
Repo size
755 MB
Likes
0
Public
Click a slice to open those files.
.bin378 MB · 100%
From the Hugging Face model README
This model is trained on the PSST Challenge data, with a subset of TIMIT that was augmented using Room Impulse Response (RIR). A file containing the list of TIMIT IDs is in the repository (timit-ids.txt)
The model was finetuned on Wav2vec 2.0 Base, No finetuning, and the results on the validation set were PER: 21.8%, FER: 9.6%.