← Search

Afshin Oroojlooy

3 accepted papers

2026

Tree-based Dialogue Reinforced Policy Optimization for Red-Teaming Attacks

ICLR 2026poster

Despite recent rapid progress in AI safety, current large language models remain vulnerable to adversarial attacks in multi-turn interaction settings, where attackers strategically adapt their prompts across conversation turns and pose a more critical yet realistic challenge. Existing approaches tha…

Cited by 0SourceScholar
2020

AttendLight: Universal Attention-Based Reinforcement Learning Model for Traffic Signal Control

NeurIPS 2020poster

We propose AttendLight, an end-to-end Reinforcement Learning (RL) algorithm for the problem of traffic signal control. Previous approaches for this problem have the shortcoming that they require training for each new intersection with a different structure or traffic flow distribution. AttendLight s…

2018

Reinforcement Learning for Solving the Vehicle Routing Problem

NeurIPS 2018poster

We present an end-to-end framework for solving the Vehicle Routing Problem (VRP) using reinforcement learning. In this approach, we train a single policy model that finds near-optimal solutions for a broad range of problem instances of similar size, only by observing the reward signals and following…

Cited by 1551SourcePDFScholar