← Search

Reza Ghoddoosian

3 accepted papers

2026

MERGE: Guided Vision-Language Models for Multi-Actor Event Reasoning and Grounding in Human–Robot Interaction

ICRA 2026poster

We introduce MERGE, a system for situational grounding of actors, objects, and events in dynamic human–robot group interactions. Effective collaboration in such settings requires consistent situational awareness, built on persistent representations of people and objects and an episodic abstraction o…

2023

Weakly-Supervised Action Segmentation and Unseen Error Detection in Anomalous Instructional Videos

ICCV 2023poster

We present a novel method for weakly-supervised action segmentation and unseen error detection in anomalous instructional videos. In the absence of an appropriate dataset for this task, we introduce the Anomalous Toy Assembly (ATA) dataset, which comprises 1152 untrimmed videos of 32 participants as…

Cited by 19PDFScholar
2022

Weakly-Supervised Online Action Segmentation in Multi-View Instructional Videos

CVPR 2022poster

This paper addresses a new problem of weakly-supervised online action segmentation in instructional videos. We present a framework to segment streaming videos online at test time using Dynamic Programming and show its advantages over greedy sliding window approach. We improve our framework by introd…

Cited by 26PDFScholar