LEGO: A Lifting-Free Approach for Exocentric-to-Egocentric Video Generation

TL;DR Given a single third-person video, LEGO generates what a person in it sees, without estimating depth or building a point cloud.

Model weights

Coming soon. We are polishing the code and checkpoints and will release them here.

Citation

@article{cho2026lego,
  title   = {LEGO: A Lifting-Free Approach for Exocentric-to-Egocentric Video Generation},
  author  = {Cho, Suhwan and Choi, Yonwoo and Kim, Soongjin and Park, Jicheol and Lim, Taegyu},
  journal = {arXiv preprint arXiv:2610.12442},
  year    = {2026}
}
Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Paper for suhwan-cho/lego