Downloads · 30 days
376
58% of all-time downloads
Fleurrr/A2World-World-Model
A2World-World-Model is a image-to-video model from Fleurrr. Use it for the image-to-video task on the model card, and read the license before you ship it in a product. It is set up for cosmos. The card lists the license as other.
A2World is a multi-view action-conditioned diffusion world model trained to predict future robot observations from initial camera observations and a 20-step action chunk. A2World-sim adds pose-guided visual history an…
Downloads · 30 days
376
58% of all-time downloads
All-time downloads
649
Public
Repo size
10.1 GB
Likes
0
Public
Click a slice to open those files.
.pt10.1 GB · 100%
From the Hugging Face model README
A2World is a multi-view action-conditioned diffusion world model trained to predict future robot observations from initial camera observations and a 20-step action chunk. A2World-sim adds pose-guided visual history and autoregressive rollout support.
a2world-pretrained.pt: heterogeneous robot-data pretrained A2World checkpoint.a2world-libero.pt: history-aware A2World-sim checkpoint adapted on LIBERO.Both released files use EMA weights, the standard net.* inference prefix, and BF16 storage compatible with the public rollout pipeline.
The checkpoints are derivative models of nvidia/Cosmos-Predict2-2B-Video2World and are governed by the NVIDIA Open Model License included in this repository. Source code is separately available under Apache-2.0.
@inproceedings{huang2026a2world,
title={Learning Transferable Dynamics Priors from Action to World Modeling},
author={Huang, Ze and Zhang, Jiahui and Liu, Hairuo and Zhang, Chenxi and Cheng, Ran and Zhang, Li},
booktitle={Proceedings of the European Conference on Computer Vision (ECCV)},
year={2026},
}