← Search

Yair Zick

9 accepted papers

2026

Hedonic Neurons: A Mechanistic Mapping of Latent Coalitions in Transformer MLPs

ICLR 2026poster

Fine-tuned Large Language Models (LLMs) encode rich task-specific features, but the form of these representations—especially within MLP layers—remains unclear. Empirical inspection of LoRA updates shows that new features concentrate in mid-layer MLPs, yet the scale of these layers obscures meaningfu…

Cited by 0SourceScholar
2025

Heterogeneous Multi-Agent Bandits with Parsimonious Hints

AAAI 2025technical

We study a hinted heterogeneous multi-agent multi-armed bandits problem (HMA2B), where agents can query low-cost observations (hints) in addition to pulling arms. In this framework, each of the M agents has a unique reward distribution over K arms, and in T rounds, they can observe the reward of the…

Cited by 0SourcePDFScholar
2024

Axiomatic Aggregations of Abductive Explanations

AAAI 2024technical

The recent criticisms of the robustness of post hoc model approximation explanation methods (like LIME and SHAP) have led to the rise of model-precise abductive explanations. For each data point, abductive explanations provide a minimal subset of features that are sufficient to generate the outcome.…

2024

Fair and Welfare-Efficient Constrained Multi-Matchings under Uncertainty

NeurIPS 2024poster

We study fair allocation of constrained resources, where a market designer optimizes overall welfare while maintaining group fairness. In many large-scale settings, utilities are not known in advance, but are instead observed after realizing the allocation. We therefore estimate agent utilities usin…

2023

Percentile Criterion Optimization in Offline Reinforcement Learning

NeurIPS 2023poster

In reinforcement learning, robust policies for high-stakes decision-making problems with limited data are usually computed by optimizing the percentile criterion. The percentile criterion is optimized by constructing an uncertainty set that contains the true model with high probability and optimizin…