2026
Evaluating Text Creativity across Diverse Domains: a Dataset and Large Language Model Evaluator
ICLR 2026poster
Creativity evaluation remains a challenging frontier for large language models (LLMs). Current evaluations heavily rely on inefficient and costly human judgments, hindering progress in enhancing machine creativity. While automated methods exist, ranging from psychological testing to heuristic- or pr…