← Search

Saurabh Kumar Pandey

7 accepted papers

2026

Measuring Meta-Cultural Competency: A Spectral Framework for LLM Knowledge Structures

ICML 2026poster

Most existing cultural evaluation frameworks for large language models (LLMs) focus on matching model outputs to ground-truth answers, primarily measuring factual cultural awareness. This overlooks whether models internalize broader cultural structure and pluralism. We introduce a spectral-analysis-…

Cited by 0SourceScholar
2025

CULTURALLY YOURS: A Reading Assistant for Cross-Cultural Content

COLING 2025system demonstrations

Users from diverse cultural backgrounds frequently face challenges in understanding content from various online sources that are written by people from a different culture. This paper presents CULTURALLY YOURS (CY), a first-of-its-kind cultural reading assistant tool designed to identify culture-spe…

2025

Meta-Cultural Competence: Climbing the Right Hill of Cultural Awareness

NAACL 2025long

Numerous recent studies have shown that Large Language Models (LLMs) are biased towards a Western and Anglo-centric worldview, which compromises their usefulness in non-Western cultural settings. However, “culture” is a complex, multifaceted topic, and its awareness, representation, and modeling in…

Cited by 0SourcePDFScholar
2025

Reading between the Lines: Can LLMs Identify Cross-Cultural Communication Gaps?

NAACL 2025long

In a rapidly globalizing and digital world, content such as book and product reviews created by people from diverse cultures are read and consumed by others from different corners of the world. In this paper, we investigate the extent and patterns of gaps in understandability of book reviews due to…

2025

SMAB: MAB based word Sensitivity Estimation Framework and its Applications in Adversarial Text Generation

NAACL 2025long

To understand the complexity of sequence classification tasks, Hahn et al. (2021) proposed sensitivity as the number of disjoint subsets of the input sequence that can each be individually changed to change the output. Though effective, calculating sensitivity at scale using this framework is costly…

2024

Evaluating ChatGPT against Functionality Tests for Hate Speech Detection

COLING 2024main

Large language models like ChatGPT have recently shown a great promise in performing several tasks, including hate speech detection. However, it is crucial to comprehend the limitations of these models to build robust hate speech detection systems. To bridge this gap, our study aims to evaluate the…

Cited by 4SourcePDFScholar
2023

CONTRASTE: Supervised Contrastive Pre-training With Aspect-based Prompts For Aspect Sentiment Triplet Extraction

EMNLP 2023long findings

Existing works on Aspect Sentiment Triplet Extraction (ASTE) explicitly focus on developing more efficient fine-tuning techniques for the task. Instead, our motivation is to come up with a generic approach that can improve the downstream performances of multiple ABSA tasks simultaneously. Towards th…

Cited by 0SourcecodeScholar