Downloads · 30 days
32
15% of all-time downloads
PhanithLIM/whisper-tiny-khmer-ct2
whisper-tiny-khmer-ct2 is a automatic speech recognition model from PhanithLIM. Use it when you need speech turned into text. It is set up for transformers. The card lists the license as apache-2.0.
PhanithLIM/whisper-tiny-aug-19-april-lightning-v1 is a fine-tuned version of OpenAI's Whisper ASR model adapted specifically for the Khmer language. Built on the small variant of Whisper and optimized using FasterWhis…
Downloads · 30 days
32
15% of all-time downloads
All-time downloads
211
Public
Repo size
75.5 MB
Likes
1
Public
Click a slice to open those files.
.bin75.5 MB · 98%
From the Hugging Face model README
PhanithLIM/whisper-tiny-aug-19-april-lightning-v1 is a fine-tuned version of OpenAI's Whisper ASR model adapted specifically for the Khmer language. Built on the small variant of Whisper and optimized using FasterWhisper, this model provides efficient and accurate speech-to-text transcription for Khmer audio.
pip install faster-whisper
from faster_whisper import WhisperModel
# Load the model
model = WhisperModel("PhanithLIM/whisper-tiny-khmer-ct2", compute_type="int8", local_files_only=False, beam_size=5)
# Transcribe Khmer audio
segments, info = model.transcribe("your_audio_file.wav")
# Print segments
for segment in segments:
print(f"{segment.start:.2f}s --> {segment.end:.2f}s: {segment.text}")
This model can be integrated into real-time systems using tools such as:
CTranslate2 is a fast inference engine for transformer models, optimized for CPU and GPU deployment, especially in production environments. It's developed by the team behind OpenNMT, and it's widely used in speech and machine translation systems, including FasterWhisper, which is a CTranslate2 port of OpenAI’s Whisper.