← Search

Yasunori Ohishi

4 accepted papers

2023

Masked Modeling Duo: Learning Representations by Encouraging Both Networks to Model the Input

ICASSP 2023accepted

Masked Autoencoders is a simple yet powerful self-supervised learning method. However, it learns representations indirectly by reconstructing masked input patches. Several methods learn representations directly by predicting representations of masked patches; however, we think using all patches to e…

Cited by 0SourceScholar
2022

Echo-Aware Adaptation of Sound Event Localization and Detection in Unknown Environments

ICASSP 2022accepted

Our goal is to develop a sound event localization and detection (SELD) system that works robustly in unknown environments. A SELD system trained on known environment data is degraded in an unknown environment due to environmental effects such as reverberation and noise not contained in the training…

Cited by 0SourceScholar
2022

Multi-View And Multi-Modal Event Detection Utilizing Transformer-Based Multi-Sensor Fusion

ICASSP 2022accepted

We tackle a challenging task: multi-view and multi-modal event detection that detects events in a wide-range real environment by utilizing data from distributed cameras and microphones and their weak labels. In this task, distributed sensors are utilized complementarily to capture events that are di…

Cited by 0SourceScholar
2020

Trilingual Semantic Embeddings of Visually Grounded Speech with Self-Attention Mechanisms

ICASSP 2020accepted

We propose a trilingual semantic embedding model that associates visual objects in images with segments of speech signals corresponding to spoken words in an unsupervised manner. Unlike the existing models, our model incorporates three different languages, namely, English, Hindi, and Japanese. To bu…

Cited by 0SourceScholar