← Search

Jin Wen

1 accepted papers

2026

On the Evaluation of Capability Estimation Methods for Large Language Models

AAAI 2026technical

The emergence of large language models (LLMs) marks a transformative era in artificial intelligence~(AI). However, systematically evaluating the capability of LLMs is challenging due to the necessity of a large number of labeled test data. To tackle this problem, in the conventional AI field, AutoEv

Cited by 0SourcePDFScholar