← Search

James Allan

5 accepted papers

2026

Hedonic Neurons: A Mechanistic Mapping of Latent Coalitions in Transformer MLPs

ICLR 2026poster

Fine-tuned Large Language Models (LLMs) encode rich task-specific features, but the form of these representations—especially within MLP layers—remains unclear. Empirical inspection of LoRA updates shows that new features concentrate in mid-layer MLPs, yet the scale of these layers obscures meaningfu…

Cited by 0SourceScholar
2024

Discovering Biases in Information Retrieval Models Using Relevance Thesaurus as Global Explanation

EMNLP 2024main

Most of the efforts in interpreting neural relevance models have been on local explanations, which explain the relevance of a document to a query. However, local explanations are not effective in predicting the model’s behavior on unseen texts. We aim at explaining a neural relevance model by provid…

Cited by 0SourcePDFScholar
2024

Language Concept Erasure for Language-invariant Dense Retrieval

EMNLP 2024main

Multilingual models aim for language-invariant representations but still prominently encode language identity. This, along with the scarcity of high-quality parallel retrieval data, limits their performance in retrieval. We introduce LANCER, a multi-task learning framework that improves language-inv…

Cited by 1SourcePDFScholar