Actor-18M is the largest human-centric video dataset (1.6M videos / 18M images) constructed to capture view-invariant identity:
@inproceedings{guo2026wildactor,
title={WildActor: Unconstrained Identity-Preserving Video Generation},
author={Qin Guo and Tianyu Yang and Xuanhua He and Fei Shen and Yong Zhang and Zhuoliang Kang and Xiaoming Wei and Dan Xu},
booktitle={Forty-third International Conference on Machine Learning},
year={2026},
url={https://openreview.net/forum?id=wXkCkP8TtK}
}