Downloads · 30 days
0
ashishkblink/NuralVoiceSTT
NuralVoiceSTT is a automatic speech recognition model from ashishkblink. Use it when you need speech turned into text. The card lists the license as apache-2.0.
Downloads · 30 days
0
Access
Public
Updated Jan 7, 2026
Repo size
208 MB
Likes
0
Public
Click a slice to open those files.
.fst114 MB · 53%
From the Hugging Face model README
Developed by Blink Digital
Accurate universal English speech-to-text model optimized for both callcenter and wideband audio scenarios.
NuralVoiceSTT is a state-of-the-art English (US) speech recognition model developed by Blink Digital. This model delivers high-accuracy speech-to-text conversion with support for both narrowband (callcenter) and wideband audio formats.
am/ - Acoustic model filesconf/ - Configuration files (MFCC, model config)graph/ - Language model graph files (FST files, phones, words)ivector/ - I-vector files for speaker adaptationimport json
import wave
from vosk import Model, KaldiRecognizer, SetLogLevel
# Set log level
SetLogLevel(0)
# Load the NuralVoiceSTT model
model = Model("path/to/NuralVoiceSTT")
# Process audio file
wf = wave.open("audio.wav", "rb")
rec = KaldiRecognizer(model, wf.getframerate())
rec.SetWords(True)
# Recognize speech
while True:
data = wf.readframes(4000)
if len(data) == 0:
break
if rec.AcceptWaveform(data):
result = json.loads(rec.Result())
print(result["text"])
# Get final result
result = json.loads(rec.FinalResult())
print(result["text"])
from huggingface_hub import snapshot_download
model_path = snapshot_download(
repo_id="ashishkblink/NuralVoiceSTT",
local_dir="./models/NuralVoiceSTT"
)
pip install vosk
NuralVoiceSTT is optimized for:
This model is distributed under the Apache License 2.0, which allows:
If you use this model, please cite:
@misc{nuralvoicestt2024,
title={NuralVoiceSTT: High-Accuracy English Speech-to-Text Model},
author={Blink Digital},
year={2024},
publisher={Blink Digital},
url={https://huggingface.co/ashishkblink/NuralVoiceSTT}
}
NuralVoiceSTT is developed and maintained by Blink Digital, a leading provider of AI-powered speech recognition solutions. For more information, visit our Hugging Face profile.
For questions, issues, or commercial inquiries, please contact Blink Digital through our Hugging Face profile.