2025
Language Models Are Capable of Metacognitive Monitoring and Control of Their Internal Activations
NeurIPS 2025poster
Large language models (LLMs) can sometimes report the strategies they actually use to solve tasks, yet at other times seem unable to recognize those strategies that govern their behavior. This suggests a limited degree of metacognition --- the capacity to monitor one's own cognitive processes for su…