← Search

Rajasimman Madhivanan

4 accepted papers

2025

ET-Former: Efficient Triplane Deformable Attention for 3D Semantic Scene Completion From Monocular Camera

IROS 2025

We introduce ET-Former, a novel end-to-end algorithm for semantic scene completion using a single monocular camera. Our approach generates a semantic occupancy map from single RGB observation while simultaneously providing uncertainty estimates for semantic predictions. By designing a triplane-based

Cited by 3SourcecodeScholar
2025

HomeEmergency - Using Audio to Find and Respond to Emergencies in the Home

RA-L 2025

In the United States alone accidental home deaths exceed 128,000 per year. Our work aims to enable home robots who respond to emergency scenarios in the home, preventing injuries and deaths. We introduce a new dataset of household emergencies based in the ThreeDWorld simulator. Each scenario in our

Cited by 0SourceScholar
2024

VLPG-Nav: Object Navigation Using Visual Language Pose Graph and Object Localization Probability Maps

IROS 2024poster

We present VLPG-Nav, a visual language navigation method for guiding robots to specified objects within household scenes. Unlike existing methods primarily focused on navigating the robot toward objects, our approach considers the additional challenge of centering the object within the robot’s camer…

Cited by 1SourceScholar
2023

Learning to View: Decision Transformers for Active Object Detection

ICRA 2023poster

Active perception describes a broad class of techniques that couple planning and perception systems to move the robot in a way to give the robot more information about the environment. In most robotic systems, perception is typically independent of motion planning. For example, traditional object de…

Cited by 18SourceScholar