Downloads · 30 days
7
13% of all-time downloads
Sparkplugx1904/whisper-tiny-id
whisper-tiny-id is a automatic speech recognition model from Sparkplugx1904. Use it when you need speech turned into text.
This model is a fine-tuned version of openai/whisper-tiny for Automatic Speech Recognition (ASR) in Indonesian (id). It supports transcription of Indonesian speech into text across various audio conditions, with perfo…
Downloads · 30 days
7
13% of all-time downloads
All-time downloads
55
Public
Parameters
37.8M
151 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors151 MB · 99%
From the Hugging Face model README
This model is a fine-tuned version of openai/whisper-tiny for Automatic Speech Recognition (ASR) in Indonesian (id).
It supports transcription of Indonesian speech into text across various audio conditions, with performance and resource usage depending on the selected model size.
Model variants (tiny, base, small, medium, large) differ in accuracy, speed, and hardware requirements. Users should select the size that best matches their constraints and objectives.
This model was fine-tuned using Mozilla Common Voice v23.0 (Indonesian).
Common Voice is a publicly available, community-driven speech dataset released by Mozilla under a permissive license.
Dataset characteristics such as speaker diversity, recording quality, and utterance length may influence model behavior.
The model is typically evaluated using Word Error Rate (WER).
Evaluation results may vary depending on dataset, domain, audio conditions, and model size.
| Step | Training Loss |
|---|---|
| 100 | 1.282900 |
| 200 | 0.682300 |
| 300 | 0.568900 |
| 400 | 0.487500 |
| 500 | 0.372700 |
| 600 | 0.375500 |
| 700 | 0.276200 |
| 800 | 0.226000 |
| 900 | 0.223800 |
| 1000 | 0.188600 |
| 1100 | 0.164300 |
| 1200 | 0.151400 |
| 1300 | 0.130000 |
| 1400 | 0.133900 |
| 1500 | 0.119700 |
| 1550 | 0.117300 |