← Search

Jonathan S. Ilgen

1 accepted papers

2024

MediQ: Question-Asking LLMs and a Benchmark for Reliable Interactive Clinical Reasoning

NeurIPS 2024poster

Users typically engage with LLMs interactively, yet most existing benchmarks evaluate them in a static, single-turn format, posing reliability concerns in interactive scenarios. We identify a key obstacle towards reliability: LLMs are trained to answer any question, even with incomplete context or i…

Cited by 16SourcePDFScholar