← Search

Behzad Mehrbakhsh

1 accepted papers

2025

Contamination Budget: Trade-offs Between Breadth, Depth and Difficulty

IJCAI 2025

Contamination in large language models (LLMs), and machine learning more broadly, refers to the inclusion of equal --or very similar-- examples in both training and test sets. This phenomenon usually translates into better test performance. Here we explore when this contamination is performed intent