← Search

Pinyan Lu

14 accepted papers

2026

A Provable Expressiveness Hierarchy in Hybrid Linear-Full Attention

ICML 2026poster

Transformers serve as the foundation of most modern large language models. To mitigate the quadratic complexity of standard full attention, various efficient attention mechanisms, such as linear and hybrid attention, have been developed. A fundamental gap remains: their expressive power relative to …

Cited by 0SourceScholar
2026

Muon in Associative Memory Learning: Training Dynamics and Scaling Laws

ICML 2026poster

Muon updates matrix parameters via the matrix sign of the gradient and has shown strong empirical gains, yet its dynamics and scaling behavior remain unclear in theory. We study Muon in a linear associative memory model with softmax retrieval and a hierarchical frequency spectrum over query–answer p…

Cited by 0SourceScholar
2026

Understanding the Ability of LLMs to Handle Character-Level Perturbation

ICML 2026poster

This work investigates the resilience of contemporary large language models (LLMs) against frequent character-level perturbations. We examine three types of character-level perturbations including introducing numerous typos within words, shuffling the characters in each word, and inserting a large n…

Cited by 0SourceScholar
2025

Bandit Learning in Matching Markets with Indifference

ICLR 2025poster

A rich line of recent works studies how participants in matching markets learn their unknown preferences through iterative interactions with each other. The two sides of participants in the market can be respectively formulated as players and arms in the bandit problem. To ensure market stability, t…

Cited by 0SourcePDFScholar
2025

Incentives for Early Arrival in Cooperative Games (Extended Abstract)

IJCAI 2025

We study cooperative games where players join sequentially, and the value generated by those who have joined at any point must be irrevocably divided among these players. We introduce two desiderata for the value division mechanism: that the players should have incentives to join as early as possibl

Cited by 0SourcePDFScholar
2023

Revocable Deep Reinforcement Learning with Affinity Regularization for Outlier-Robust Graph Matching

ICLR 2023poster

Graph matching (GM) has been a building block in various areas including computer vision and pattern recognition. Despite recent impressive progress, existing deep GM methods often have obvious difficulty in handling outliers, which are ubiquitous in practice. We propose a deep reinforcement learnin…

Cited by 11SourcePDFScholar
2021

Online Selection Problems against Constrained Adversary

ICML 2021spotlight

Inspired by a recent line of work in online algorithms with predictions, we study the constrained adversary model that utilizes predictions from a different perspective. Prior works mostly focused on designing simultaneously robust and consistent algorithms, without making assumptions on the quality…

Cited by 19SourcePDFScholar
2020

Strategyproof Mechanism for Two Heterogeneous Facilities with Constant Approximation Ratio

IJCAI 2020poster

In this paper, we study the two-facility location game with optional preference where the acceptable set of facilities for each agent could be different and an agent's cost is his distance to the closest facility within his acceptable set. The objective is to minimize the total cost of all agents wh…

Cited by 0SourcePDFScholar
2016

Combinatorial Multi-Armed Bandit with General Reward Functions

NeurIPS 2016poster

In this paper, we study the stochastic combinatorial multi-armed bandit (CMAB) framework that allows a general nonlinear reward function, whose expected value may not depend only on the means of the input random variables but possibly on the entire distributions of these variables. Our framework ena…

Cited by 175SourcePDFScholar