← Search

Bikash Dutta

3 accepted papers

2026

ACID Test: A Benchmark for Cultural Safety and Alignment in LALMs

AAAI 2026technical

Large Audio Language Models (LALMs) are transforming AI by processing and generating human language directly from audio. As these models proliferate in real-world applications, it becomes critical to evaluate their performance to ensure equitable and safe use across diverse linguistic and cultural c

Cited by 0SourcePDFScholar
2025

Can RAG-Driven Enhancements Amplify Audio LLMs for Low-Resource Languages?

ICASSP 2025accepted

The proliferation of Large Language Models (LLMs) has transformed Natural Language Processing (NLP), yet their development has largely overlooked low-resource languages. This paper addresses this disparity by evaluating three prominent Large Audio Language Models (LALMs) – LTU-AS, GAMA, and Pengi –…

Cited by 0SourceScholar
2024

BirdCollect: A Comprehensive Benchmark for Analyzing Dense Bird Flock Attributes

AAAI 2024technical

Automatic recognition of bird behavior from long-term, un controlled outdoor imagery can contribute to conservation efforts by enabling large-scale monitoring of bird populations. Current techniques in AI-based wildlife monitoring have focused on short-term tracking and monitoring birds individually…

Cited by 2SourcePDFScholar