← Search

Sina Mokhtarzadeh Azar

3 accepted papers

2026

EgoControl: Controllable Egocentric Video Generation via 3D Full-Body Poses

CVPR 2026

Egocentric video generation with fine-grained control through body motion is a key requirement towards embodied AI agents that can simulate, predict, and plan actions. In this work, we propose EgoControl, a pose-controllable video diffusion model trained on egocentric data. We train a video predicti

Cited by 0SourcecodeScholar
2025

SyncVP: Joint Diffusion for Synchronous Multi-Modal Video Prediction

CVPR 2025poster

Predicting future video frames is essential for decision-making systems, yet RGB frames alone often lack the information needed to fully capture the underlying complexities of the real world. To address this limitation, we propose a multi-modal framework for Synchronous Video Prediction (SyncVP) tha…

2019

Convolutional Relational Machine for Group Activity Recognition

CVPR 2019poster

We present an end-to-end deep Convolutional Neural Network called Convolutional Relational Machine (CRM) for recognizing group activities that utilizes the information in spatial relations between individual persons in image or video. It learns to produce an intermediate spatial representation (acti…

Cited by 155PDFScholar