Downloads · 30 days
20
1% of all-time downloads
versae/whisper-large-v3
whisper-large-v3 is a automatic speech recognition model from versae. Use it when you need speech turned into text. It is set up for transformers. The card lists the license as apache-2.0.
Version 3 of OpenAI's Whisper Large model converted from https://openaipublic.azureedge.net/main/whisper/models/e5b1a55b89c1367dacf97e3e19bfd829a01529dbfdeefa8caeb59b3f1b81dadb/large-v3.pt using HF's conversion script.
Downloads · 30 days
20
1% of all-time downloads
All-time downloads
1.8K
Public
Parameters
1.5B
12.3 GB on disk
Likes
18
Public
Click a slice to open those files.
.bin6.2 GB · 50%
From the Hugging Face model README
Version 3 of OpenAI's Whisper Large model converted from https://openaipublic.azureedge.net/main/whisper/models/e5b1a55b89c1367dacf97e3e19bfd829a01529dbfdeefa8caeb59b3f1b81dadb/large-v3.pt using HF's conversion script.
Whisper is a pre-trained model for automatic speech recognition (ASR) and speech translation. Trained on 680k hours of labelled data, Whisper models demonstrate a strong ability to generalise to many datasets and domains without the need for fine-tuning.
Whisper was proposed in the paper Robust Speech Recognition via Large-Scale Weak Supervision by Alec Radford et al. from OpenAI. The original code repository can be found here.