Downloads · 30 days
0
scdrand23/ReVQom
ReVQom is a machine learning model from scdrand23. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as mit.
Pretrained checkpoints for ReVQom, a learned feature codec for multi-agent collaborative perception. ReVQom compresses BEV features via a 1x1 bottleneck and multi-stage residual vector quantization, transmitting only…
Downloads · 30 days
0
Access
Public
Updated Aug 9, 2026
Repo size
68.2 MB
Likes
0
Public
Click a slice to open those files.
.pth34.1 MB · 100%
From the Hugging Face model README
Pretrained checkpoints for ReVQom, a learned feature codec for multi-agent collaborative perception. ReVQom compresses BEV features via a 1x1 bottleneck and multi-stage residual vector quantization, transmitting only per-pixel code indices (6-30 bits per pixel, 273x-1365x compression vs raw features).
| File | Config | Dataset | [email protected]/[email protected] |
|---|---|---|---|
revqom_s_k64_dairv2x/net_epoch30.pth | ReVQom-S: K=64, n_q=3, C_rr=16, EMA 0.8 (18 bpp, 455x) | DAIR-V2X | 0.751/0.646 |
The checkpoint directory includes the exact training config.yaml, and the evaluation matches the paper's experiment records. Additional checkpoints (ReVQom-M, K=256) will be added after re-evaluation.
pip install -U huggingface_hub
hf download scdrand23/ReVQom --local-dir checkpoints
python revqom/tools/inference.py --model_dir checkpoints/revqom_s_k64_dairv2x --fusion_method intermediate
See the GitHub repository for installation and dataset preparation.
@inproceedings{shenkut2026revqom,
title={Residual Vector Quantization for Communication-Efficient Multi-Agent Perception},
author={Shenkut, Dereje and Kumar, B.V.K. Vijaya},
booktitle={IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)},
year={2026}
}