Downloads · 30 days
13
11% of all-time downloads
jaman21/faster-whisper-large-v3
faster-whisper-large-v3 is a automatic speech recognition model from jaman21. Use it when you need speech turned into text. It is set up for ctranslate2. The card lists the license as mit.
This repository contains the conversion of openai/whisper-large-v3 to the CTranslate2 model format.
Downloads · 30 days
13
11% of all-time downloads
All-time downloads
118
Public
Repo size
3.1 GB
Likes
0
Public
Click a slice to open those files.
.bin3.1 GB · 100%
From the Hugging Face model README
This repository contains the conversion of openai/whisper-large-v3 to the CTranslate2 model format.
This model can be used in CTranslate2 or projects based on CTranslate2 such as faster-whisper.
from faster_whisper import WhisperModel
model = WhisperModel("large-v3")
segments, info = model.transcribe("audio.mp3")
for segment in segments:
print("[%.2fs -> %.2fs] %s" % (segment.start, segment.end, segment.text))
The original model was converted with the following command:
ct2-transformers-converter --model openai/whisper-large-v3 --output_dir faster-whisper-large-v3 \
--copy_files tokenizer.json preprocessor_config.json --quantization float16
Note that the model weights are saved in FP16. This type can be changed when the model is loaded using the compute_type option in CTranslate2.
For more information about the original model, see its model card.