2026
When Silence Is Golden: Can LLMs Learn to Abstain in Temporal QA and Beyond?
ICLR 2026poster
Large language models (LLMs) rarely admit uncertainty, often producing fluent but misleading answers, rather than abstaining (i.e., refusing to answer). This weakness is even evident in temporal question answering (QA), where models frequently ignore time-sensitive evidence and conflate facts across…