← Search

Dick Botteldooren

8 accepted papers

2026

NAACA: Training-Free NeuroAuditory Attentive Cognitive Architecture with Oscillatory Working Memory for Salience-Driven Attention Gating

ICML 2026poster

Audio provides critical situational cues, yet current Audio Language Models (ALMs) face an attention bottleneck in long-form recordings where dominant background patterns can dilute rare, salient events. We introduce NAACA, a training-free NeuroAuditory Attentive Cognitive Architecture that reframes…

Cited by 0SourceScholar
2025

Sound-Based Recognition of Touch Gestures and Emotions for Enhanced Human-Robot Interaction

ICASSP 2025accepted

Emotion recognition and touch gesture decoding are crucial for advancing human-robot interaction (HRI), especially in social environments where emotional cues and tactile perception play important roles. However, many humanoid robots, such as Pepper, Nao, and Furhat, lack full-body tactile skin, lim…

Cited by 0SourceScholar
2024

Multi-Level Graph Learning For Audio Event Classification And Human-Perceived Annoyance Rating Prediction

ICASSP 2024accepted

WHO’s report on environmental noise estimates that 22 M people suffer from chronic annoyance related to noise caused by audio events (AEs) from various sources. Annoyance may lead to health issues and adverse effects on metabolic and cognitive systems. In cities, monitoring noise levels does not pro…

Cited by 0SourceScholar
2024

No More Mumbles: Enhancing Robot Intelligibility Through Speech Adaptation

RA-L 2024

Spoken language interaction is at the heart of interpersonal communication, and people flexibly adapt their speech to different individuals and environments. It is surprising that robots, and by extension other digital devices, are not equipped to adapt their speech and instead rely on fixed speech

Cited by 7SourcecodeScholar
2023

Adaptive Axonal Delays in Feedforward Spiking Neural Networks for Accurate Spoken Word Recognition

ICASSP 2023accepted

Spiking neural networks (SNN) are a promising research avenue for building accurate and efficient automatic speech recognition systems. Recent advances in audio-to-spike encoding and training algorithms enable SNN to be applied in practical tasks. Biologically-inspired SNN communicates using sparse…

Cited by 0SourceScholar
2022

Axonal Delay as a Short-Term Memory for Feed Forward Deep Spiking Neural Networks

ICASSP 2022accepted

The information of spiking neural networks (SNNs) are propagated between the adjacent biological neuron by spikes, which provides a computing paradigm with the promise of simulating the human brain. Recent studies have found that the time delay of neurons plays an important role in the learning proc…

Cited by 0SourceScholar
2021

Rule-Embedded Network for Audio-Visual Voice Activity Detection in Live Musical Video Streams

ICASSP 2021accepted

Detecting anchor’s voice in live musical streams is an important preprocessing step for music and speech signal processing. Existing approaches to voice activity detection (VAD) primarily rely on audio, however, audio-based VAD is difficult to effectively focus on the target voice in noisy environme…

Cited by 0SourceScholar