Downloads · 30 days
0
inference4j/silero-vad
silero-vad is a voice activity detection model from inference4j. Use it for the voice activity detection task on the model card, and read the license before you ship it in a product. It is set up for onnx. The card lists the license as mit.
ONNX export of Silero VAD, a lightweight and fast voice activity detection model. Detects speech segments in audio with high accuracy and low latency.
Downloads · 30 days
0
Access
Public
Updated Feb 13, 2026
Repo size
2.3 MB
Likes
0
Public
Click a slice to open those files.
.onnx2.3 MB · 100%
From the Hugging Face model README
ONNX export of Silero VAD, a lightweight and fast voice activity detection model. Detects speech segments in audio with high accuracy and low latency.
Mirrored for use with inference4j, an inference-only AI library for Java.
try (SileroVAD vad = SileroVAD.fromPretrained("models/silero-vad")) {
List<VoiceSegment> segments = vad.detect(Path.of("meeting.wav"));
for (VoiceSegment segment : segments) {
System.out.printf("Speech: %.2fs - %.2fs%n", segment.start(), segment.end());
}
}
| Property | Value |
|---|---|
| Architecture | Silero VAD (lightweight CNN + LSTM) |
| Task | Voice activity detection |
| Input | 16kHz mono audio (float32 waveform, 512-sample chunks) |
| Output | Speech probability per chunk |
| Model size | ~2 MB |
| Original source | snakers4/silero-vad |
This model is licensed under the MIT License. Original model by Silero Team.