2024
Developing a Pragmatic Benchmark for Assessing Korean Legal Language Understanding in Large Language Models
EMNLP 2024finding
Large language models (LLMs) have demonstrated remarkable performance in the legal domain, with GPT-4 even passing the Uniform Bar Exam in the U.S. However their efficacy remains limited for non-standardized tasks and tasks in languages other than English. This underscores the need for careful evalu…