← Search

Scott L Fleming

2 accepted papers

2024

MedAlign: A Clinician-Generated Dataset for Instruction Following with Electronic Medical Records

AAAI 2024technical

The ability of large language models (LLMs) to follow natural language instructions with human-level fluency suggests many opportunities in healthcare to reduce administrative burden and improve quality of care. However, evaluating LLMs on realistic text generation tasks for healthcare remains chall…

Cited by 67SourcePDFScholar
2021

Reinforcement Learning with State Observation Costs in Action-Contingent Noiselessly Observable Markov Decision Processes

NeurIPS 2021poster

Many real-world problems that require making optimal sequences of decisions under uncertainty involve costs when the agent wishes to obtain information about its environment. We design and analyze algorithms for reinforcement learning (RL) in Action-Contingent Noiselessly Observable MDPs (ACNO-MDPs)…