← Search

Yunmeng Li

2 accepted papers

2025

MQM-Chat: Multidimensional Quality Metrics for Chat Translation

COLING 2025main

The complexities of chats, such as the stylized contents specific to source segments and dialogue consistency, pose significant challenges for machine translation. Recognizing the need for a precise evaluation metric to address the issues associated with chat translation, this study introduces Multi…

2025

Rubrik’s Cube: Testing a New Rubric for Evaluating Explanations on the CUBE dataset

ACL 2025long

The performance and usability of Large-Language Models (LLMs) are driving their use in explanation generation tasks. However, despite their widespread adoption, LLM explanations have been found to be unreliable, making it difficult for users to distinguish good from bad explanations. To address this…