2025
RADAR: Benchmarking Language Models on Imperfect Tabular Data
NeurIPS 2025poster
Language models (LMs) are increasingly being deployed to perform autonomous data analyses. However, their data awareness—the ability to recognize, reason over, and appropriately handle data artifacts such as missing values, outliers, and logical inconsistencies—remains underexplored. These artifacts…