Downloads · 30 days
30
51% of all-time downloads
victorlfdev/bs-roformer-multi-q8
bs-roformer-multi-q8 is a machine learning model from victorlfdev. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.
A quantized GGUF version of the BS-Roformer model, optimized for C++ inference via Fork-BSRoformer.cpp. Supports multi-stem separation (drums, bass, vocals, others).
Downloads · 30 days
30
51% of all-time downloads
All-time downloads
59
Public
Repo size
188 MB
Likes
0
Public
Click a slice to open those files.
.gguf188 MB · 100%
From the Hugging Face model README
A quantized GGUF version of the BS-Roformer model, optimized for C++ inference via Fork-BSRoformer.cpp. Supports multi-stem separation (drums, bass, vocals, others).
BS-Roformer (Band Split RoFormer) is a neural network architecture for music source separation. It splits the audio frequency spectrum into multiple bands and processes each band independently using a transformer encoder with self-attention mechanisms. The model was originally derived from Suno AI's Bark project (text-to-music generation), where it was used internally for music source separation.
This checkpoint has been quantized to Q8 format using the GGUF library for efficient inference in C++ environments, reducing memory usage while maintaining high separation quality.
This model was trained using the framework from Music Source Separation Training, which is a PyTorch-based training framework for music source separation models. The training data consists of publicly available music datasets used by the community training efforts documented in that repository.
| Credit | Link |
|---|---|
| Suno AI | Original creators of the BS-Roformer / Mel-Band-Roformer architecture via the Bark project |
| ZFTurbo (Vladislav Sukachov) | Music Source Separation Training framework and community model training |
| anvuew | Trained BS-RoFormer checkpoint (SDR 12.45) |
| GaboxR67 | Mel-Band-Roformer checkpoints |
| 沉默の金 (chenmozhijin) | Fork-BSRoformer.cpp — C++ GGUF inference engine |
| ggerganov | GGML library for efficient tensor computation |
| dr_libs | Lightweight audio decoding library |
Download the compiled binary and run:
./bs_roformer-cli -m bs-roformer-multi-q8.gguf -a input.wav -o output.wav
See Fork-BSRoformer.cpp (https://github.com/victorlfdev/Fork-BSRoformer.cpp) for full CLI options and usage.
Via Python
from bs_roformer_cpp_cli import BsRoformerCppCLI
cli = BsRoformerCppCLI(model_path="./bs-roformer-multi-q8.gguf", device="cuda")
cli.process("input.wav", "output.wav")
Model Architecture
- Type: Band Split RoFormer (transformer-based music source separator)
- Quantization: Q8 (8-bit uniform quantization via GGUF)
- Stems: 4 (drums, bass, vocals, other)
- Input: Mono/stereo WAV audio (any sample rate, resampled internally)
- Output: 4-channel separated stems (WAV format)
License
This model is shared for research and educational purposes. The underlying BS-Roformer architecture and training methodology are derived from community efforts referenced above. Redistribution of trained weights should comply with the original training data licenses.
Acknowledgements
- ggerganov/ggml (https://github.com/ggerganov/ggml) — Efficient tensor library
- ZFTurbo/Music-Source-Separation-Training (https://github.com/ZFTurbo/Music-Source-Separation-Training) — PyTorch reference implementation
- dr_libs (https://github.com/mackron/dr_libs) — Lightweight audio library