← Search

Rakesh Kumar

3 accepted papers

2026

GeoSURGE: Geo-localization using Semantic Fusion with Hierarchy of Geographic Embeddings

CVPR 2026

Worldwide visual geo-localization aims to determine the geographic location of an image anywhere on Earth using only its visual content. Despite recent progress, learning expressive representations of geographic space remains challenging due to the inherently low-dimensional nature of geographic coo

Cited by 0SourceScholar
2021

MaAST: Map Attention with Semantic Transformers for Efficient Visual Navigation

ICRA 2021poster

Visual navigation for autonomous agents is a core task in the fields of computer vision and robotics. Learning-based methods, such as deep reinforcement learning, have the potential to outperform the classical solutions developed for this task; however, they come at a significantly increased computa…

Cited by 24SourceScholar
2017

Fast human segmentation using color and depth

ICASSP 2017accepted

Accurate segmentation of humans from live videos is an important problem to be solved in developing immersive video experience. We propose to extract the human segmentation information from color and depth cues in a video using multiple modeling techniques. The prior information from human skeleton…

Cited by 0SourceScholar