2024
BUST: Benchmark for the evaluation of detectors of LLM-Generated Text
NAACL 2024long
We introduce BUST, a comprehensive benchmark designed to evaluate detectors of texts generated by instruction-tuned large language models (LLMs). Unlike previous benchmarks, our focus lies on evaluating the performance of detector systems, acknowledging the inevitable influence of the underlying tas…