← Search

Li Ji-An

2 accepted papers

2025

Language Models Are Capable of Metacognitive Monitoring and Control of Their Internal Activations

NeurIPS 2025poster

Large language models (LLMs) can sometimes report the strategies they actually use to solve tasks, yet at other times seem unable to recognize those strategies that govern their behavior. This suggests a limited degree of metacognition --- the capacity to monitor one's own cognitive processes for su…

Cited by 0SourceScholar
2024

Linking In-context Learning in Transformers to Human Episodic Memory

NeurIPS 2024poster

Understanding connections between artificial and biological intelligent systems can reveal fundamental principles of general intelligence. While many artificial intelligence models have a neuroscience counterpart, such connections are largely missing in Transformer models and the self-attention mech…