← Search

Vigneshwaran Subbaraju

4 accepted papers

2025

Ges3ViG : Incorporating Pointing Gestures into Language-Based 3D Visual Grounding for Embodied Reference Understanding

CVPR 2025poster

3-Dimensional Embodied Reference Understanding (3DERU) combines a language description and an accompanying pointing gesture to identify the most relevant target object in a 3D scene. Although prior work has explored pure language-based 3D grounding, there has been limited exploration of 3D-ERU, whic…

2024

CAS: Fusing DNN Optimization & Adaptive Sensing for Energy-Efficient Multi-Modal Inference

RA-L 2024

Intelligent virtual agents are used to accomplish complex multi-modal tasks such as human instruction comprehension in mixed-reality environments by increasingly adopting richer, energy-intensive sensors and processing pipelines. In such applications, the <italic xmlns:mml="http://www.w3.org/1998/Ma

Cited by 0SourceScholar
2022

COSM2IC: Optimizing Real-Time Multi-Modal Instruction Comprehension

RA-L 2022

Supporting real-time, on-device execution of <italic xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">multi-modal referring instruction comprehension</i> models is an important challenge to be tackled in embodied Human-Robot Interaction. However, state-of-the

Cited by 11SourceScholar
2021

Predicting Event Memorability from Contextual Visual Semantics

NeurIPS 2021poster

Episodic event memory is a key component of human cognition. Predicting event memorability,i.e., to what extent an event is recalled, is a tough challenge in memory research and has profound implications for artificial intelligence. In this study, we investigate factors that affect event memorabilit…