Downloads · 30 days
22
1% of all-time downloads
Alvenir/wav2vec2-base-da
wav2vec2-base-da is a machine learning model from Alvenir. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.
This wav2vec2-base model has been pretrained on ~1300 hours of danish speech data. The pretraining data consists of podcasts and audiobooks and is unfortunately not public available. However, we were allowed to distri…
Downloads · 30 days
22
1% of all-time downloads
All-time downloads
3.8K
Public
Repo size
760 MB
Likes
6
Public
Click a slice to open those files.
.bin380 MB · 100%
From the Hugging Face model README
This wav2vec2-base model has been pretrained on ~1300 hours of danish speech data. The pretraining data consists of podcasts and audiobooks and is unfortunately not public available. However, we were allowed to distribute the pretrained model.
This model was pretrained on 16kHz sampled speech audio. When using the model, make sure to use speech audio sampled at 16kHz.
The pre-training was done using the fairseq library in January 2021.
It needs to be fine-tuned to perform speech recognition.
In order to finetune the model to speech recognition, you can draw inspiration from this notebook tutorial or this blog post tutorial.