← Search

Augustin Kelava

1 accepted papers

2026

Position: Stop evaluating AI with human tests, develop principled, AI-specific tests instead

ICML 2026poster

Large Language Models (LLMs) have achieved remarkable results on a range of standardized tests originally designed to assess human cognitive and psychological traits, such as intelligence and personality. While these results are often interpreted as strong evidence of human-like characteristics in L…

Cited by 0SourceScholar