← Search

Joe Eappen

5 accepted papers

2025

FLoRA: A Framework for Learning Scoring Rules in Autonomous Driving Planning Systems

RA-L 2025

In autonomous driving systems, motion planning is commonly implemented as a two-stage process: first, a trajectory proposer generates multiple candidate trajectories, then a scoring mechanism selects the most suitable trajectory for execution. For this critical selection stage, rule-based scoring me

Cited by 1SourceScholar
2024

Co-learning Planning and Control Policies Constrained by Differentiable Logic Specifications

ICRA 2024poster

Synthesizing planning and control policies in robotics is a fundamental task, further complicated by factors such as complex logic specifications and high-dimensional robot dynamics. This paper presents a novel reinforcement learning approach to solving high-dimensional robot navigation tasks with c…

Cited by 1SourceScholar
2024

Information-Directed Pessimism for Offline Reinforcement Learning

ICML 2024poster

Policy optimization from batch data, i.e., offline reinforcement learning (RL) is important when collecting data from a current policy is not possible. This setting incurs distribution mismatch between batch training data and trajectories from the current policy. Pessimistic offsets estimate mismatc…

Cited by 1SourcePDFScholar
2024

Scaling Safe Multi-Agent Control for Signal Temporal Logic Specifications

CoRL 2024poster

Existing methods for safe multi-agent control using logic specifications like Signal Temporal Logic (STL) often face scalability issues. This is because they rely either on single-agent perspectives or on Mixed Integer Linear Programming (MILP)-based planners, which are complex to optimize. These me…

Cited by 1SourcecodeScholar
2022

Model-free Neural Lyapunov Control for Safe Robot Navigation

IROS 2022poster

Model-free Deep Reinforcement Learning (DRL) controllers have demonstrated promising results on various challenging non-linear control tasks. While a model-free DRL algorithm can solve unknown dynamics and high-dimensional problems, it lacks safety assurance. Although safety constraints can be encod…

Cited by 8SourcecodeScholar