← Search

Hosana Kamiyama

3 accepted papers

2021

Age-VOX-Celeb: Multi-Modal Corpus for Facial and Speech Estimation

ICASSP 2021accepted

Estimating a speaker’s age from their speech is more challenging than age estimation from their face because of insufficiently available public corpora. To tackle this problem, we construct a new audio-visual age corpus named AgeVoxCeleb by annotating age labels to VoxCeleb2 videos. AgeVoxCeleb is t…

Cited by 0SourceScholar
2020

Improving Speaker-Attribute Estimation by Voting Based on Speaker Cluster Information

ICASSP 2020accepted

This paper proposes a general post-processing method for improving speaker-attribute estimation. Estimating speaker-specific attributes such as age and gender is an important task with a wide range of applications. While the recent proposed deep neural network-based end-to-end approach achieves high…

Cited by 0SourceScholar
2018

Soft-Target Training with Ambiguous Emotional Utterances for DNN-Based Speech Emotion Classification

ICASSP 2018accepted

This paper presents a novel emotion classification method for natural speech. One of the problems in the state-of-the-art method based on Deep Neural Network (DNN) is the paucity of the training data compared to model complexity. To solve this problem, this paper utilizes the ambiguous emotional utt…

Cited by 0SourceScholar