← Search

Shigeo Morishima

6 accepted papers

2025

Memory-Maze: Scenario Driven Visual Language Navigation Benchmark for Guiding Blind People

RA-L 2025

Visual Language Navigation (VLN) powered robots have the potential to guide blind people by understanding route instructions provided by sighted passersby. This capability allows robots to operate in environments often unknown a prior. Existing VLN models are insufficient for the scenario of navigat

Cited by 3SourceScholar
2022

Geometric Features Informed Multi-Person Human-Object Interaction Recognition in Videos

ECCV 2022poster

"Human-Object Interaction (HOI) recognition in videos is important for analyzing human activity. Most existing work focusing on visual features usually suffer from occlusion in the real-world scenarios. Such a problem will be further complicated when multiple people and objects are involved in HOIs.…

2021

Pitch-Timbre Disentanglement Of Musical Instrument Sounds Based On Vae-Based Metric Learning

ICASSP 2021accepted

This paper describes a representation learning method for disentangling an arbitrary musical instrument sound into latent pitch and timbre representations. Although such pitch-timbre disentanglement has been achieved with a variational autoencoder (VAE), especially for a predefined set of musical in…

Cited by 0SourceScholar
2019

PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human Digitization

ICCV 2019poster

We introduce Pixel-aligned Implicit Function (PIFu), an implicit representation that locally aligns pixels of 2D images with the global context of their corresponding 3D object. Using PIFu, we propose an end-to-end deep learning method for digitizing highly detailed clothed humans that can infer bot…

Cited by 1438PDFScholar