2025
Imitation Beyond Expectation Using Pluralistic Stochastic Dominance
NeurIPS 2025spotlight
Imitation learning seeks policies reflecting the values of demonstrated behaviors. Prevalent approaches learn to match or exceed the demonstrator's performance in expectation without knowing the demonstrator’s reward function. Unfortunately, this does not induce pluralistic imitators that learn to s…