Downloads · 30 days
342
8% of all-time downloads
lucadellalib/focalcodec_25hz
focalcodec_25hz is a audio-to-audio model from lucadellalib. Use it for the audio-to-audio task on the model card, and read the license before you ship it in a product. It is set up for pytorch. The card lists the license as apache-2.0.
A low-bitrate single-codebook 16 / 24 kHz speech codec based on focal modulation.
Downloads · 30 days
342
8% of all-time downloads
All-time downloads
4.3K
Public
Parameters
144M
577 MB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors577 MB · 100%
How the weights are stored.
F32144M · 100%
From the Hugging Face model README
A low-bitrate single-codebook 16 / 24 kHz speech codec based on focal modulation.
This repository contains the 25 Hz checkpoint trained on LibriTTS 960, as described in the preprints.
📜 Preprints:
🌐 Project Page: https://lucadellalib.github.io/focalcodec-web/
See the readme at: https://github.com/lucadellalib/focalcodec
@article{dellalibera2025focalcodec,
title = {{FocalCodec}: Low-Bitrate Speech Coding via Focal Modulation Networks},
author = {Luca {Della Libera} and Francesco Paissan and Cem Subakan and Mirco Ravanelli},
journal = {arXiv preprint arXiv:2502.04465},
year = {2025},
}
@article{dellalibera2025focalcodecstream,
title = {{FocalCodec-Stream}: Streaming Low-Bitrate Speech Coding via Causal Distillation},
author = {Luca {Della Libera} and Cem Subakan and Mirco Ravanelli},
journal = {arXiv preprint arXiv:2509.16195},
year = {2025},
}