← Search

Mohammad Tavakoli

2 accepted papers

2026

Beyond a Million Tokens: Benchmarking and Enhancing Long-Term Memory in LLMs

ICLR 2026poster

Evaluating the abilities of large language models (LLMs) for tasks that require long-term memory and thus long-context reasoning, for example in conversational settings, is hampered by the existing benchmarks, which often lack narrative coherence, cover narrow domains, and only test simple recall-or…

Cited by 0SourcecodeScholar
2025

Semi-Automated Construction of Sense-Annotated Datasets for Practically Any Language

COLING 2025main

High-quality sense-annotated datasets are vital for evaluating and comparing WSD systems. We present a novel approach to creating parallel sense-annotated datasets, which can be applied to any language that English can be translated into. The method incorporates machine translation, word alignment,…