2024
The Music Maestro or The Musically Challenged, A Massive Music Evaluation Benchmark for Large Language Models
ACL 2024findings
Benchmark plays a pivotal role in assessing the advancements of large language models (LLMs). While numerous benchmarks have been proposed to evaluate LLMs’ capabilities, there is a notable absence of a dedicated benchmark for assessing their musical abilities. To address this gap, we present ZIQI-E…