← Search

Yu-Min Tseng

2 accepted papers

2026

SealQA: Raising the Bar for Reasoning in Search-Augmented Language Models

ICLR 2026poster

We introduce SealQA, a challenge benchmark for evaluating SEarch-Augmented Language models on fact-seeking questions where web search yields conflicting, noisy, or unhelpful results. SealQA comes in three flavors: (1) SEAL-0 (main) and (2) SEAL-HARD, both of which assess factual accuracy and reasoni…

Cited by 0SourceScholar
2024

Two Tales of Persona in LLMs: A Survey of Role-Playing and Personalization

EMNLP 2024finding

The concept of *persona*, originally adopted in dialogue literature, has re-surged as a promising framework for tailoring large language models (LLMs) to specific context (*e.g.*, personalized search, LLM-as-a-judge). However, the growing research on leveraging persona in LLMs is relatively disorgan…