Downloads · 30 days
0
todteera/audio-restore
audio-restore is a audio-to-audio model from todteera. Use it for the audio-to-audio task on the model card, and read the license before you ship it in a product. It is set up for pytorch. The card lists the license as cc-by-nc-4.0.
Repairs speech recordings damaged by clippings. A generative model reconstructs the samples that were destroyed.
Downloads · 30 days
0
Access
Public
Updated Aug 11, 2026
Repo size
196 MB
Likes
0
Public
Click a slice to open those files.
.pt196 MB · 100%
From the Hugging Face model README
Repairs speech recordings damaged by clippings. A generative model reconstructs the samples that were destroyed.
git clone https://github.com/tdstt22/audio-restore
cd audio-restore && uv sync
python restore.py recording.wav restored.wav
| File | Description |
|---|---|
model.pt | 2D U-Net, 48.9M params — flow-matching velocity predictor over complex STFT. EMA weights plus architecture config |
The model works on spectrograms at 24 kHz mono. It is given the damaged audio and a mask marking which samples the clipping destroyed, so it never has to guess where the damage is. Rather than emitting audio directly, it predicts the direction from noise toward clean speech, and sampling follows that direction over to arrive at the reconstruction.