← Search

Yihong Guo

5 accepted papers

2026

CorrectionPlanner: Self-Correction Planner with Reinforcement Learning in Autonomous Driving

ICML 2026poster

Autonomous driving requires safe planning, but most learning-based planners lack explicit self-correction ability: once an unsafe action is proposed, there is no mechanism to correct it. Thus, we propose CorrectionPlanner, an autoregressive planner with self-correction that models planning as motion…

Cited by 0SourceScholar
2025

TSD-SR: One-Step Diffusion with Target Score Distillation for Real-World Image Super-Resolution

CVPR 2025poster

Pre-trained text-to-image diffusion models are increasingly applied to real-world image super-resolution (Real-ISR) task. Given the iterative refinement nature of diffusion models, most existing approaches are computationally expensive. While methods such as SinSR and OSEDiff have emerged to condens…

2024

Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation

NeurIPS 2024poster

Training a policy in a source domain for deployment in the target domain under a dynamics shift can be challenging, often resulting in performance degradation. Previous work tackles this challenge by training on the source domain with modified rewards derived by matching distributions between the so…

2023

Distributionally Robust Policy Gradient for Offline Contextual Bandits

AISTATS 2023poster

Learning an optimal policy from offline data is notoriously challenging, which requires the evaluation of the learning policy using data pre-collected from a static logging policy. We study the policy optimization problem in offline contextual bandits using policy gradient methods. We employ a distr…