← Search

Qingzhuo Wang

2 accepted papers

2026

A Unified Approach to Interpreting Knowledge Distillation for Large Language Models via Interactions

ICML 2026poster

Despite the success of knowledge distillation (KD) in Large Language Models (LLMs), the underlying mechanism behind its efficacy remains unclear. In this paper, we propose a unified approach to explore the common mechanism of various KD methods using interactions. Specifically, we decompose the outp…

Cited by 0SourceScholar
2026

Evaluating and Explaining Prompt Sensitivity of LLMs Using Interactions

ICML 2026poster

The remarkable capabilities of large language models (LLMs) are often undermined by their instability. Even subtle and semantically irrelevant changes in prompts can cause dramatic fluctuations in performance, a phenomenon known as prompt sensitivity. Previous studies typically evaluate prompt sensi…

Cited by 0SourceScholar