← Search

Noam Koenigstein

18 accepted papers

2026

ConEx: Human-Interpretable Saliency Maps via Concept-Aware Attribution

ICML 2026poster

Many visual explanation methods in computer vision highlight pixel importance but struggle to link these low-level cues to semantically meaningful concepts, limiting their interpretability and trustworthiness. We introduce Concept-based Explanations (ConEx), a novel framework that bridges saliency v…

Cited by 0SourceScholar
2026

Concept-Guided Fine-Tuning: Steering ViTs away from Spurious Correlations to Improve Robustness

CVPR 2026

Vision Transformers (ViTs) often degrade under distribution shifts because they rely on spurious correlations, such as background cues, rather than semantically meaningful features. Existing regularization methods, typically relying on simple foreground-background masks, which fail to capture the fi

Cited by 0SourcecodeScholar
2026

Extracting Interaction-Aware Monosemantic Concepts in Recommender Systems

AAAI 2026technical

We present a method for extracting monosemantic neurons, defined as latent dimensions that align with coherent and interpretable concepts, from user and item embeddings in recommender systems. Our approach employs a Sparse Autoencoder (SAE) to reveal semantic structure within pretrained representati

Cited by 0SourcePDFScholar
2026

Fidelity-Aware Recommendation Explanations via Stochastic Path Integration

AAAI 2026technical

Explanation fidelity, which measures how accurately an explanation reflects a model’s true reasoning, remains critically underexplored in recommender systems. We introduce SPINRec (Stochastic Path Integration for Neural Recommender Explanations), a model-agnostic approach that adapts path-integratio

Cited by 0SourcePDFScholar
2026

Rethinking Saliency Maps: A Cognitive Human Aligned Taxonomy and Evaluation Framework for Explanations

AAAI 2026technical

Saliency maps have become a cornerstone of visual explanation in deep learning, yet there remains no consensus on their intended purpose and their alignment with specific user queries. This fundamental ambiguity undermines both the evaluation and practical utility of explanation methods. In this pap

Cited by 0SourcePDFScholar
2025

BEE: Metric-Adapted Explanations via Baseline Exploration-Exploitation

AAAI 2025technical

Two prominent challenges in explainability research involve 1) the nuanced evaluation of explanations and 2) the modeling of missing information through baseline representations. The existing literature introduces diverse evaluation metrics, each scrutinizing the quality of explanations through dist…

2025

Soft Local Completeness: Rethinking Completeness in XAI

ICCV 2025accepted

Completeness is a widely discussed property in explainability research, requiring that the attributions sum to the model's response to the input. While completeness intuitively suggests that the model's prediction is "completely explained" by the attributions, its global formulation alone is insuffi…

2024

Improving LLM Attributions with Randomized Path-Integration

EMNLP 2024finding

We present Randomized Path-Integration (RPI) - a path-integration method for explaining language models via randomization of the integration path over the attention information in the model. RPI employs integration on internal attention scores and their gradients along a randomized path, which is dy…

2024

InterrogateLLM: Zero-Resource Hallucination Detection in LLM-Generated Answers

ACL 2024long

Despite the many advances of Large Language Models (LLMs) and their unprecedented rapid evolution, their impact and integration into every facet of our daily lives is limited due to various reasons. One critical factor hindering their widespread adoption is the occurrence of hallucinations, where LL…

2024

LLM Explainability via Attributive Masking Learning

EMNLP 2024finding

In this paper, we introduce Attributive Masking Learning (AML), a method designed for explaining language model predictions by learning input masks. AML trains an attribution model to identify influential tokens in the input for a given language model’s prediction. The central concept of AML is to t…

2024

SEGLLM: Topic-Oriented Call Segmentation Via LLM-Based Conversation Synthesis

ICASSP 2024accepted

Transcriptions of phone calls are of significant value across diverse fields, such as sales, customer service, healthcare, and law enforcement. Nevertheless, the analysis of these recorded conversations can be an arduous and time-intensive process, especially when dealing with long and multifaceted…

Cited by 0SourceScholar
2024

Unsupervised Topic-Conditional Extractive Summarization

ICASSP 2024accepted

Summarization techniques strive to create a concise summary that conveys the essential information from a given document. However, these techniques are often inadequate for summarizing longer documents containing multiple pages of semantically complex content with various topics. Hence, in this work…

Cited by 0SourceScholar
2023

Efficient Discovery and Effective Evaluation of Visual Perceptual Similarity: A Benchmark and Beyond

ICCV 2023poster

Visual similarities discovery (VSD) is an important task with broad e-commerce applications. Given an image of a certain object, the goal of VSD is to retrieve images of different objects with high perceptual visual similarity. Although being a highly addressed problem, the evaluation of proposed me…

Cited by 6PDFcodeScholar
2023

Visual Explanations via Iterated Integrated Attributions

ICCV 2023poster

We introduce Iterated Integrated Attributions (IIA) - a generic method for explaining the predictions of vision models. IIA employs iterative integration across the input image, the internal representations generated by the model, and their gradients, yielding precise and focused explanation maps. W…

Cited by 24PDFcodeScholar
2022

Metricbert: Text Representation Learning Via Self-Supervised Triplet Training

ICASSP 2022accepted

We present MetricBERT, a BERT-based model that learns to embed text under a well-defined similarity metric while simultaneously adhering to the “traditional” masked-language task. We focus on downstream tasks of learning similarities for recommendations where we show that MetricBERT outperforms stat…

Cited by 0SourceScholar
2021

Cold Start Revisited: A Deep Hybrid Recommender with Cold-Warm Item Harmonization

ICASSP 2021accepted

Collaborative filtering-based recommender systems are known to suffer from the item cold-start problem. Most recent attempts to mitigate this problem presented parametric approaches, such as deep content based models. In this paper, we show that a straightforward application of parametric models may…

Cited by 0SourceScholar