Downloads · 30 days
32
100% of all-time downloads
aTrain-core/KB-WhisperSwedish
KB-WhisperSwedish is a automatic speech recognition model from aTrain-core. Use it when you need speech turned into text. It is set up for ctranslate2. The card lists the license as apache-2.0.
This repository contains only the CTranslate2 files from KBLab/kb-whisper-large, a Swedish Whisper large-v3 model created by KBLab at the National Library of Sweden. The files were copied unchanged from upstream revis…
Downloads · 30 days
32
100% of all-time downloads
All-time downloads
32
Public
Repo size
3.1 GB
Likes
0
Public
Click a slice to open those files.
.bin3.1 GB · 100%
From the Hugging Face model README
This repository contains only the CTranslate2 files from
KBLab/kb-whisper-large, a
Swedish Whisper large-v3 model created by KBLab at the National Library of
Sweden. The files were copied unchanged from upstream revision
d5d5984b4d8f7c4847a8ea203f1976285fb28300.
The upstream model card contains the training data, evaluation results, limitations, acknowledgements, and citation. KBLab reports that the model was trained on more than 50,000 hours of Swedish speech.
from faster_whisper import WhisperModel
model = WhisperModel("aTrain-core/KB-WhisperSwedish", device="cpu", compute_type="int8")
segments, info = model.transcribe("audio.mp3", language="sv", word_timestamps=True)
for segment in segments:
print(f"[{segment.start:.2f}s -> {segment.end:.2f}s] {segment.text}")
The files were tested with faster-whisper==1.2.1 on CPU using int8.
The upstream model is distributed under the Apache License 2.0. See
LICENSE. KB-Whisper is a product of KBLab at the National Library
of Sweden. Please use the citation provided in the
upstream model card.