← Search

Sandra Mitrovic

1 accepted papers

2024

BUST: Benchmark for the evaluation of detectors of LLM-Generated Text

NAACL 2024long

We introduce BUST, a comprehensive benchmark designed to evaluate detectors of texts generated by instruction-tuned large language models (LLMs). Unlike previous benchmarks, our focus lies on evaluating the performance of detector systems, acknowledging the inevitable influence of the underlying tas…