ACL 2025long0 citations

Crab: A Novel Configurable Role-Playing LLM with Assessing Benchmark

Kai He, Yucheng Huang, Wenqing Wang, Delong Ran, Dongming Sheng, Junxuan Huang, Qika Lin, Jiaxing Xu

Abstract

This study introduces Crab, a novel Configurable Role-Playing (RP) LLM with Assessing Benchmark, which consists of Role-Centric Dataset Curation, Persona-Embodying LLM Construction, and Comprehensive Benchmark Creation for RP dialogue generation. Distinct from traditional RP models that employ only several preset roles, Crab enables dynamic configuration of desired roles, thereby enhancing related flexibility and adaptability. To effectively train RP-LLMs, we curated the largest RP training dataset. The dataset provides a detailed role overview for each dialogue, including character profile, conversation scenario, and tagged topic, capturing a broad range of role-based behaviors, emotions, and interactions. We also noticed that current benchmarks lack both proper evaluation standards and methods. Thus, to validate RP-LLMs’ effectiveness, we introduced a new benchmark containing an evaluation standard, a test dataset with manual annotations, and a reward model RoleRM designed to automatically assess specific aspects of RP while aligning with human perception. Sufficient experiments reveal that RoleRM significantly outperforms ChatGPT and other evaluation methods in conducting fine-grained evaluations of RP. Also, RP-LLMs powered by Crab demonstrate superior performance across various fine-grained aspects.

BibTeX
@inproceedings{he-etal-2025-crab,
    title = "Crab: A Novel Configurable Role-Playing {LLM} with Assessing Benchmark",
    author = "He, Kai  and
      Huang, Yucheng  and
      Wang, Wenqing  and
      Ran, Delong  and
      Sheng, Dongming  and
      Huang, Junxuan  and
      Lin, Qika  and
      Xu, Jiaxing  and
      Liu, Wenqiang  and
      Feng, Mengling",
    editor = "Che, Wanxiang  and
      Nabende, Joyce  and
      Shutova, Ekaterina  and
      Pilehvar, Mohammad Taher",
    booktitle = "Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)",
    month = jul,
    year = "2025",
    address = "Vienna, Austria",
    publisher = "Association for Computational Linguistics",
    url = "https://aclanthology.org/2025.acl-long.731/",
    doi = "10.18653/v1/2025.acl-long.731",
    pages = "15030--15052",
    ISBN = "979-8-89176-251-0"
}
Crab: A Novel Configurable Role-Playing LLM with Assessing Benchmark · ACL 2025