← Search

Dinesh Tewari

3 accepted papers

2026

CURVE: A Benchmark for Cultural and Multilingual Long Video Reasoning

CVPR 2026

Recent advancements in video models have shown tremendous progress, particularly in long video understanding. However, current benchmarks predominantly feature western-centric data and English as the dominant language, introducing significant biases in evaluation. To address this, we introduce CURVE

Cited by 0SourceScholar
2024

IndicGenBench: A Multilingual Benchmark to Evaluate Generation Capabilities of LLMs on Indic Languages

ACL 2024long

As large language models (LLMs) see increasing adoption across the globe, it is imperative for LLMs to be representative of the linguistic diversity of the world. India is a linguistically diverse country of 1.4 Billion people. To facilitate research on multilingual LLM evaluation, we release IndicG…

2023

Building Socio-culturally Inclusive Stereotype Resources with Community Engagement

NeurIPS 2023poster

With rapid development and deployment of generative language models in global settings, there is an urgent need to also scale our measurements of harm, not just in the number and types of harms covered, but also how well they account for local cultural contexts, including marginalized identities and…

Cited by 21SourcePDFScholar