Downloads · 30 days
214
11% of all-time downloads
NaiveNeuron/whisper-large-v3-sk
whisper-large-v3-sk is a machine learning model from NaiveNeuron. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as mit.
This model is a fine-tuned version of openai/whisper-large-v3. It is adapted for Slovak ASR using SloPalSpeech: 2,806 hours of aligned, ≤30 s speech–text pairs from official plenary sessions of the Slovak National Cou…
Downloads · 30 days
214
11% of all-time downloads
All-time downloads
1.9K
Public
Parameters
1.5B
6.2 GB on disk
Likes
4
Public
Click a slice to open those files.
.safetensors6.2 GB · 100%
From the Hugging Face model README
This model is a fine-tuned version of openai/whisper-large-v3.
It is adapted for Slovak ASR using SloPalSpeech: 2,806 hours of aligned, ≤30 s speech–text pairs from official plenary sessions of the Slovak National Council.
| Dataset | Base WER | Fine-tuned WER | Δ (abs) |
|---|---|---|---|
| Common Voice 21 (sk) | 20.8 | 11.6 | -9.2 |
| FLEURS (sk) | 9.2 | 5.5 | -3.7 |
Numbers from the paper’s final benchmark runs.
1e-5 with weight decay 0.01 to prevent overfittingFor more details, please see our paper on arXiv. If you use this model in your work, please cite it as:
@misc{božík2025slopalspeech2800hourslovakspeech,
title={SloPalSpeech: A 2,800-Hour Slovak Speech Corpus from Parliamentary Data},
author={Erik Božík and Marek Šuppa},
year={2025},
eprint={2509.19270},
archivePrefix={arXiv},
primaryClass={cs.CL},
url={https://arxiv.org/abs/2509.19270},
}
This work was supported by VÚB Banka who provided the GPU resources and backing necessary to accomplish it, enabling progress in Slovak ASR research.