ACL 2025long0 citations

CEAES: Bidirectional Reinforcement Learning Optimization for Consistent and Explainable Essay Assessment

Xia Li, Wenjing Pan

Abstract

Most current automated essay quality assessment systems treat score prediction and feedback generation as separate tasks, overlooking the fact that scores provide a quantitative evaluation of quality, while feedback offers a qualitative assessment. Both aspects reflect essay quality from different perspectives, and they are inherently consistent and can reinforce each other. In this paper, we propose a novel bidirectional reinforcement learning framework that effectively utilizes this consistency constraint to jointly optimize score prediction and feedback generation, ensuring mutual reinforcement and alignment between them. In this way, our model is hope to obtain a simultaneous accurate ratings and consistent text feedback. We conducted extensive experiments on publicly available datasets. The results demonstrate that our approach surpasses the current state-of-the-art models, enhancing both scoring accuracy and feedback quality.

BibTeX
@inproceedings{li-pan-2025-ceaes,
    title = "{CEAES}: Bidirectional Reinforcement Learning Optimization for Consistent and Explainable Essay Assessment",
    author = "Li, Xia  and
      Pan, Wenjing",
    editor = "Che, Wanxiang  and
      Nabende, Joyce  and
      Shutova, Ekaterina  and
      Pilehvar, Mohammad Taher",
    booktitle = "Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)",
    month = jul,
    year = "2025",
    address = "Vienna, Austria",
    publisher = "Association for Computational Linguistics",
    url = "https://aclanthology.org/2025.acl-long.1273/",
    doi = "10.18653/v1/2025.acl-long.1273",
    pages = "26267--26279",
    ISBN = "979-8-89176-251-0"
}