← Search

Archan Misra

6 accepted papers

2026

NeuroLiDAR: Adaptive Frame Rate Depth Sensing Via Neuromorphic Event-LiDAR Fusion

ICRA 2026poster

LiDARs are widely used for 3D depth reconstruction, but their performance is often limited by inherent hardware constraints that impose trade-offs between range, spatial resolution, and frame rate. Many LiDAR systems typically operate at low frame rates (e.g., 5-10 Hz), prioritizing long-range sensi…

2025

Ges3ViG : Incorporating Pointing Gestures into Language-Based 3D Visual Grounding for Embodied Reference Understanding

CVPR 2025poster

3-Dimensional Embodied Reference Understanding (3DERU) combines a language description and an accompanying pointing gesture to identify the most relevant target object in a 3D scene. Although prior work has explored pure language-based 3D grounding, there has been limited exploration of 3D-ERU, whic…

2024

CAS: Fusing DNN Optimization & Adaptive Sensing for Energy-Efficient Multi-Modal Inference

RA-L 2024

Intelligent virtual agents are used to accomplish complex multi-modal tasks such as human instruction comprehension in mixed-reality environments by increasingly adopting richer, energy-intensive sensors and processing pipelines. In such applications, the <italic xmlns:mml="http://www.w3.org/1998/Ma

Cited by 0SourceScholar
2024

D2SR: Decentralized Detection, De-Synchronization, and Recovery of LiDAR Interference

IROS 2024poster

We address the challenge of multi-LiDAR interference, an issue of growing importance as LiDAR sensors are embedded in a growing set of pervasive devices. We introduce a novel approach named D2SR, enabling decentralized interference detection, mitigation, and recovery without explicit coordination am…

Cited by 0SourceScholar
2024

EyeGraph: Modularity-aware Spatio Temporal Graph Clustering for Continuous Event-based Eye Tracking

NeurIPS 2024poster

Continuous tracking of eye movement dynamics plays a significant role in developing a broad spectrum of human-centered applications, such as cognitive skills (visual attention and working memory) modeling, human-machine interaction, biometric user authentication, and foveated rendering. Recently neu…

Cited by 1SourcePDFScholar
2022

COSM2IC: Optimizing Real-Time Multi-Modal Instruction Comprehension

RA-L 2022

Supporting real-time, on-device execution of <italic xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">multi-modal referring instruction comprehension</i> models is an important challenge to be tackled in embodied Human-Robot Interaction. However, state-of-the

Cited by 11SourceScholar