← Search

Sui Wenjuan

1 accepted papers

2025

Diagnosing Failures in Large Language Models’ Answers: Integrating Error Attribution into Evaluation Framework

ACL 2025finding

With the widespread application of Large Language Models (LLMs) in various tasks, the mainstream LLM platforms generate massive user-model interactions daily. In order to efficiently analyze the performance of models and diagnose failures in their answers, it is essential to develop an automated fra…