arXiv · 2510.17731
Can Image-To-Video Models Simulate Pedestrian Dynamics?
Abstract
Recent high-performing image-to-video (I2V) models based on variants of the diffusion transformer (DiT) have displayed remarkable inherent world-modeling capabilities by virtue of training on large scale video datasets. We investigate whether these models can generate realistic pedestrian movement patterns in crowded public scenes. Our framework conditions I2V models on keyframes extracted from pedestrian trajectory benchmarks, then evaluates their trajectory prediction performance using quantitative measures of pedestrian dynamics.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Aaron Appelle, Jerome P. Lynch. 2025-10-20. Can Image-To-Video Models Simulate Pedestrian Dynamics?. https://arxiv.org/abs/2510.17731
Cite the original work for its findings. Save a collection to share your selection of sources.