← Search

Wei-Di Chang

5 accepted papers

2026

The Surprising Difficulty of Search in Model-Based Reinforcement Learning

ICML 2026poster

This paper investigates search in model-based reinforcement learning (RL). Conventional wisdom holds that long-term predictions and compounding errors are the primary obstacles for model-based RL. We challenge this view, showing that search is not a plug-and-play replacement for a learned policy. Su…

Cited by 4SourceScholar
2025

Generalizable Imitation Learning Through Pre-Trained Representations

ICRA 2025

In this paper, we leverage self-supervised vision transformer models and their emergent semantic abilities to improve the generalization abilities of imitation learning policies. We introduce DVK, an imitation learning algorithm that leverages rich pre-trained Visual Transformer patch-level embeddin

Cited by 5SourceScholar
2023

For SALE: State-Action Representation Learning for Deep Reinforcement Learning

NeurIPS 2023poster

In reinforcement learning (RL), representation learning is a proven tool for complex image-based tasks, but is often overlooked for environments with low-level states, such as physical control problems. This paper introduces SALE, a novel approach for learning embeddings that model the nuanced inte…

2020

One-Shot Informed Robotic Visual Search in the Wild

IROS 2020poster

We consider the task of underwater robot navigation for the purpose of collecting scientifically relevant video data for environmental monitoring. The majority of field robots that currently perform monitoring tasks in unstructured natural environments navigate via path-tracking a pre-specified sequ…

Cited by 16SourcecodeScholar
2017

Underwater multi-robot convoying using visual tracking by detection

IROS 2017poster

We present a robust multi-robot convoying approach that relies on visual detection of the leading agent, thus enabling target following in unstructured 3-D environments. Our method is based on the idea of tracking-by-detection, which interleaves efficient model-based object detection with temporal f…

Cited by 81SourcecodeScholar