← Search

Carrie Ye

2 accepted papers

2026

Beyond a Million Tokens: Benchmarking and Enhancing Long-Term Memory in LLMs

ICLR 2026poster

Evaluating the abilities of large language models (LLMs) for tasks that require long-term memory and thus long-context reasoning, for example in conversational settings, is hampered by the existing benchmarks, which often lack narrative coherence, cover narrow domains, and only test simple recall-or…

Cited by 0SourcecodeScholar
2025

Not What the Doctor Ordered: Surveying LLM-based De-identification and Quantifying Clinical Information Loss

EMNLP 2025

De-identification in the healthcare setting is an application of NLP where automated algorithms are used to remove personally identifying information of patients (and, sometimes, providers). With the recent rise of generative large language models (LLMs), there has been a corresponding rise in the n