← Search

sriram srinivasan

9 accepted papers

2026

Self-Refining Vision Language Model for Robotic Failure Detection and Reasoning

ICLR 2026poster

Reasoning about failures is crucial for building reliable and trustworthy robotic systems. Prior approaches either treat failure reasoning as a closed-set classification problem or assume access to ample human annotations. Failures in the real world are typically subtle, combinatorial, and difficult…

Cited by 0SourceScholar
2023

SCA: Streaming Cross-Attention Alignment For Echo Cancellation

ICASSP 2023accepted

End-to-End deep learning has shown promising results for speech enhancement tasks, such as noise suppression, dereverberation, and speech separation. However, most state-of-the-art methods for echo cancellation are either classical DSP-based or hybrid DSP-ML algorithms. Components such as the delay…

Cited by 0SourceScholar
2021

ICASSP 2021 Acoustic Echo Cancellation Challenge: Datasets, Testing Framework, and Results

ICASSP 2021accepted

The ICASSP 2021 Acoustic Echo Cancellation Challenge is intended to stimulate research in the area of acoustic echo cancellation (AEC), which is an important part of speech enhancement and still a top issue in audio communication and conferencing systems. Many recent AEC studies report good performa…

Cited by 0SourceScholar
2021

ICASSP 2021 Deep Noise Suppression Challenge

ICASSP 2021accepted

The Deep Noise Suppression (DNS) challenge is designed to foster innovation in the area of noise suppression to achieve superior perceptual speech quality. We recently organized a DNS challenge special session at INTERSPEECH 2020 where we open-sourced training and test datasets for researchers to tr…

Cited by 0SourceScholar
2021

Interactive Speech and Noise Modeling for Speech Enhancement

AAAI 2021technical

Speech enhancement is challenging because of the diversity of background noise types. Most of the existing methods are focused on modelling the speech rather than the noise. In this paper, we propose a novel idea to model speech and noise simultaneously in a two-branch convolutional neural network,…

2019

Two-temperature logistic regression based on the Tsallis divergence

AISTATS 2019poster

We develop a variant of multiclass logistic regression that is significantly more robust to noise. The algorithm has one weight vector per class and the surrogate loss is a function of the linear activations (one per class). The surrogate loss of an example with linear activation vector $\mathbf{a}…

Cited by 29SourcePDFScholar
2018

Actor-Critic Policy Optimization in Partially Observable Multiagent Environments

NeurIPS 2018poster

Optimization of parameterized policies for reinforcement learning (RL) is an important and challenging problem in artificial intelligence. Among the most common approaches are algorithms based on gradient ascent of a score function representing discounted return. In this paper, we examine the role o…

2017

Neural Episodic Control

ICML 2017poster

Deep reinforcement learning methods attain super-human performance in a wide range of environments. Such methods are grossly inefficient, often taking orders of magnitudes more data than humans to achieve reasonable performance. We propose Neural Episodic Control: a deep reinforcement learning agent…

Cited by 448SourcePDFScholar
2016

Unifying Count-Based Exploration and Intrinsic Motivation

NeurIPS 2016poster

We consider an agent's uncertainty about its environment and the problem of generalizing this uncertainty across states. Specifically, we focus on the problem of exploration in non-tabular reinforcement learning. Drawing inspiration from the intrinsic motivation literature, we use density models to…

Cited by 1892SourcePDFScholar