← Search

Tessa Han

4 accepted papers

2024

MedSafetyBench: Evaluating and Improving the Medical Safety of Large Language Models

NeurIPS 2024poster

As large language models (LLMs) develop increasingly sophisticated capabilities and find applications in medical settings, it becomes important to assess their medical safety due to their far-reaching implications for personal and public health, patient safety, and human rights. However, there is li…

2022

Which Explanation Should I Choose? A Function Approximation Perspective to Characterizing Post Hoc Explanations

NeurIPS 2022accept

A critical problem in the field of post hoc explainability is the lack of a common foundational goal among methods. For example, some methods are motivated by function approximation, some by game theoretic notions, and some by obtaining clean visualizations. This fragmentation of goals causes not on…