← Search

Malek Mechergui

2 accepted papers

2024

Expectation Alignment: Handling Reward Misspecification in the Presence of Expectation Mismatch

NeurIPS 2024poster

Detecting and handling misspecified objectives, such as reward functions, has been widely recognized as one of the central challenges within the domain of Artificial Intelligence (AI) safety research. However, even with the recognition of the importance of this problem, we are unaware of any works t…

2024

Goal Alignment: Re-analyzing Value Alignment Problems Using Human-Aware AI

AAAI 2024technical

While the question of misspecified objectives has gotten much attention in recent years, most works in this area primarily focus on the challenges related to the complexity of the objective specification mechanism (for example, the use of reward functions). However, the complexity of the objective s…