2026
Ego-Grounding for Personalized Question-Answering in Egocentric Videos
CVPR 2026
We present the first systematic analysis of multimodal large language models (MLLMs) in personalized question-answering requiring ego-grounding - the ability to understand the camera-wearer in egocentric videos. To this end, we introduce MyEgo, the first egocentric VideoQA dataset designed to evalua