2024
Large Language Models Are Poor Clinical Decision-Makers: A Comprehensive Benchmark
EMNLP 2024main
The adoption of large language models (LLMs) to assist clinicians has attracted remarkable attention. Existing works mainly adopt the close-ended question-answering (QA) task with answer options for evaluation. However, many clinical decisions involve answering open-ended questions without pre-set o…