2025
InnerThoughts: Disentangling Representations and Predictions in Large Language Models
AISTATS 2025poster
Large language models (LLMs) contain substantial factual knowledge which is commonly elicited by multiple-choice question-answering prompts. Internally, such models process the prompt through multiple transformer layers, building varying representations of the problem within its hidden states. Ultim…