← Search

Johanne Medina

1 accepted papers

2026

Can LLMs Detect Their Confabulations? Estimating Reliability in Uncertainty-Aware Language Models

AAAI 2026technical

Large Language Models (LLMs) are prone to generating fluent but incorrect content, known as confabulation, which poses increasing risks in multi-turn or agentic applications where outputs may be reused as context. In this work, we investigate how in-context information influences model behavior and

Cited by 0SourcePDFScholar