← Search

Reya Vir

2 accepted papers

2026

LakeQA: A Benchmark for Complex Exploratory QA over a Million-Scale Data Lake

ICML 2026poster

Recent large language models (LLMs) have shown rapid progress on reading-based question answering (QA), where the evidence is explicitly provided or trivially retrievable. In contrast, real-world questions are often not paired with accurate evidence documents. The useful evidence resides in a massiv…

Cited by 0SourceScholar
2025

PROMPTEVALS: A Dataset of Assertions and Guardrails for Custom Production Large Language Model Pipelines

NAACL 2025long

Large language models (LLMs) are increasingly deployed in specialized production data processing pipelines across diverse domains—such as finance, marketing, and e-commerce. However, when running them in production across many inputs, they often fail to follow instructions or meet developer expectat…