← Search

Akhil Agnihotri

4 accepted papers

2026

Multi-Objective Preference Optimization: Improving Human Alignment of Generative Models

ICML 2026poster

Post-training LLMs with RLHF and preference optimization methods (e.g., DPO, IPO) has greatly improved alignment, yet these approaches assume a single objective. In reality, humans express multiple, often conflicting objectives, such as helpfulness and harmlessness, with no natural scalarization. We…

Cited by 0SourceScholar
2024

e-COP : Episodic Constrained Optimization of Policies

NeurIPS 2024poster

In this paper, we present the e-COP algorithm, the first policy optimization algorithm for constrained Reinforcement Learning (RL) in episodic (finite horizon) settings. Such formulations are applicable when there are separate sets of optimization criteria and constraints on a system's behavior. We…

Cited by 0SourcePDFScholar
2022

Investigating the Impact of Multi-LiDAR Placement on Object Detection for Autonomous Driving

CVPR 2022poster

The past few years have witnessed an increasing interest in improving the perception performance of LiDARs on autonomous vehicles. While most of the existing works focus on developing new deep learning algorithms or model architectures, we study the problem from the physical design perspective, i.e.…

Cited by 63PDFcodeScholar