← Search

Helen Zhou

3 accepted papers

2025

Learning to Route LLMs with Confidence Tokens

ICML 2025poster

Large language models (LLMs) have demonstrated impressive performance on several tasks and are increasingly deployed in real-world applications. However, especially in high-stakes settings, it becomes vital to know when the output of an LLM may be unreliable. Depending on whether an answer is trustw…

Cited by 0SourcePDFScholar
2024

Timing as an Action: Learning When to Observe and Act

AISTATS 2024poster

In standard reinforcement learning setups, the agent receives observations and performs actions at evenly spaced intervals. However, in many real-world settings, observations are expensive, forcing agents to commit to courses of action for designated periods of time. Consider that doctors, after eac…

Cited by 3SourcePDFScholar