← Search

Enrico Pallotta

3 accepted papers

2026

EgoControl: Controllable Egocentric Video Generation via 3D Full-Body Poses

CVPR 2026

Egocentric video generation with fine-grained control through body motion is a key requirement towards embodied AI agents that can simulate, predict, and plan actions. In this work, we propose EgoControl, a pose-controllable video diffusion model trained on egocentric data. We train a video predicti

Cited by 0SourcecodeScholar
2025

SyncVP: Joint Diffusion for Synchronous Multi-Modal Video Prediction

CVPR 2025poster

Predicting future video frames is essential for decision-making systems, yet RGB frames alone often lack the information needed to fully capture the underlying complexities of the real world. To address this limitation, we propose a multi-modal framework for Synchronous Video Prediction (SyncVP) tha…

2024

Race Against the Machine: A Fully-Annotated, Open-Design Dataset of Autonomous and Piloted High-Speed Flight

RA-L 2024

Unmanned aerial vehicles, and multi-rotors in particular, can now perform dexterous tasks in impervious environments, from infrastructure monitoring to emergency deliveries. Autonomous drone racing has emerged as an ideal benchmark to develop and evaluate these capabilities. Its challenges include a

Cited by 14SourcecodeScholar