← Search

Xiu Jiang

2 accepted papers

2025

MaXIFE: Multilingual and Cross-lingual Instruction Following Evaluation

ACL 2025long

With the rapid adoption of large language models (LLMs) in natural language processing, the ability to follow instructions has emerged as a key metric for evaluating their practical utility. However, existing evaluation methods often focus on single-language scenarios, overlooking the challenges and…

2023

CAME: Contrastive Automated Model Evaluation

ICCV 2023poster

The Automated Model Evaluation (AutoEval) framework entertains the possibility of evaluating a trained machine learning model without resorting to a labeled testing set. Despite the promise and some decent results, the existing AutoEval methods heavily rely on computing distribution shifts between…

Cited by 9PDFcodeScholar