2025
CUPCase: Clinically Uncommon Patient Cases and Diagnoses Dataset
AAAI 2025technical
Medical benchmark datasets significantly contribute to developing Large Language Models (LLMs) for medical knowledge extraction, diagnosis, summarization, and other uses. Yet, current benchmarks are mainly derived from exam questions given to medical students or cases described in the medical litera…