← Search

Md Alimoor Reza

6 accepted papers

2025

Logic-RAG: Augmenting Large Multimodal Models with Visual-Spatial Knowledge for Road Scene Understanding

ICRA 2025

Large multimodal models (LMMs) are increasingly integrated into autonomous driving systems for user interaction. However, their limitations in fine-grained spatial reasoning pose challenges for system interpretability and user trust. We introduce Logic-RAG, a novel Retrieval-Augmented Generation (RA

Cited by 3SourcecodeScholar
2023

Few-Shot Segmentation and Semantic Segmentation for Underwater Imagery

IROS 2023poster

This paper tackles image segmentation problems for underwater environments. First, we introduce a novel under-water animal-centric dataset with dense pixel-level annotations containing diverse fine-grained animal categories to mitigate the lack of diverse categories in the existing benchmarks. Then,…

Cited by 5SourcecodeScholar
2019

Automatic Annotation for Semantic Segmentation in Indoor Scenes

IROS 2019poster

Domestic robots could eventually transform our lives, but safely operating in home environments requires a rich understanding of indoor scenes. Learning-based techniques for scene segmentation require large-scale, pixel-level annotations, which are laborious and expensive to collect. We propose an a…

Cited by 8SourceScholar
2016

RGB-D multi-view object detection with object proposals and shape context

IROS 2016poster

We propose a novel approach for multi-view object detection in 3D scenes reconstructed from RGB-D sensor. We utilize shape based representation using local shape context descriptors along with the voting strategy which is supported by unsupervised object proposals generated from 3D point cloud data.…

Cited by 6SourceScholar