2024
F-Eval: Asssessing Fundamental Abilities with Refined Evaluation Methods
ACL 2024long
Large language models (LLMs) garner significant attention for their unprecedented performance, leading to an increasing number of researches evaluating LLMs. However, these evaluation benchmarks are limited to assessing the instruction-following capabilities, overlooking the fundamental abilities th…