← Search

Zhaoyi Sun

1 accepted papers

2025

MEDEC: A Benchmark for Medical Error Detection and Correction in Clinical Notes

ACL 2025finding

Several studies have shown that Large Language Models (LLMs) can answer medical questions correctly, even outperforming the average human score in some medical exams. However, to our knowledge, no study has been conducted to assess the ability of language models to validate existing or generated med…