Downloads · 30 days
52
4% of all-time downloads
AstralZander/yoruba_ASR
yoruba_ASR is a automatic speech recognition model from AstralZander. Use it when you need speech turned into text. It is set up for transformers. The card lists the license as apache-2.0.
facebook/wav2vec2-xls-r-300m fine-tuned on openslr (SLR86) and mozilla-foundation/commonvoice120 for Yoruba language.
Downloads · 30 days
52
4% of all-time downloads
All-time downloads
1.2K
Public
Repo size
2.5 GB
Likes
6
Public
Click a slice to open those files.
.bin1.3 GB · 100%
From the Hugging Face model README
facebook/wav2vec2-xls-r-300m fine-tuned on openslr (SLR86) and mozilla-foundation/common_voice_12_0 for Yoruba language.
WER: 0.51
from huggingsound import SpeechRecognitionModel
model = SpeechRecognitionModel("AstralZander/yoruba_ASR")
audio_paths = [audio_path] # List with paths to audio
transcriptions = model.transcribe(audio_paths)
transcriptions # List of transcriptions, timestamps and probabilities
transcriptions[ind_audio]['transcription'] # Transcription of audio with the ind_audio index from the audio_paths list