Downloads · 30 days
228
9% of all-time downloads
DAVIAN-Robotics/EgoX
EgoX is a video-to-video model from DAVIAN-Robotics. Use it for the video-to-video task on the model card, and read the license before you ship it in a product. It is set up for diffusers. The card lists the license as mit.
This repository provides model weights of EgoX, a video-to-video generation model that synthesizes egocentric (first-person) videos from a single exocentric (third-person) video. EgoX is built on top of a large-scale…
Downloads · 30 days
228
9% of all-time downloads
All-time downloads
2.6K
Public
Repo size
3.4 GB
Likes
9
Public
Click a slice to open those files.
.safetensors3.4 GB · 100%
From the Hugging Face model README
This repository provides model weights of EgoX, a video-to-video generation model that synthesizes egocentric (first-person) videos from a single exocentric (third-person) video.
EgoX is built on top of a large-scale video diffusion backbone and enables exo-to-ego viewpoint transformation without requiring multi-view inputs.
For detailed results, implementation details, and demo videos, please refer to our paper and project repository.
Please refer to the Quick Start section for instructions on running inference and required preprocessing steps.
If you find this model or code useful in your research, please cite our paper:
@misc{kang2025egoxegocentricvideogeneration,
title={EgoX: Egocentric Video Generation from a Single Exocentric Video},
author={Taewoong Kang and Kinam Kim and Dohyeon Kim and Minho Park and Junha Hyung and Jaegul Choo},
year={2025},
eprint={2512.08269},
archivePrefix={arXiv},
primaryClass={cs.CV},
url={https://arxiv.org/abs/2512.08269},
}
This work builds upon the valuable open-source efforts of
4DNeX and
EgoExo4D.
We sincerely appreciate their contributions to the computer vision and robotics communities.