← Search

Carter Teplica

1 accepted papers

2025

SCIURus: Shared Circuits for Interpretable Uncertainty Representations in Language Models

NAACL 2025long

We investigate the mechanistic sources of uncertainty in large language models (LLMs), an area with important implications for language model reliability and trustworthiness. To do so, we conduct a series of experiments designed to identify whether the factuality of generated responses and a model’s…