← Search

Catherine Glossop

8 accepted papers

2026

Learning to Drive Anywhere With Model-Based Reannotation

RA-L 2026

Developing broadly generalizable visual navigation policies for robots is a significant challenge, primarily constrained by the availability of large-scale, diverse training data. While curated datasets collected by researchers offer high quality, their limited size restricts policy generalization.

Cited by 12SourcecodeScholar
2026

Learning to Drive Anywhere with Model-Based Reannotation

ICRA 2026poster

Developing broadly generalizable visual navigation policies for robots is a significant challenge, primarily constrained by the availability of large-scale, diverse training data. While curated datasets collected by researchers offer high quality, their limited size restricts policy generalization. …

2026

OmniVLA: An Omni-Modal Vision-Language-Action Model for Robot Navigation

ICRA 2026poster

Humans can flexibly interpret and compose different goal specifications, such as language instructions, spatial coordinates, or visual references, when navigating to a destination. In contrast, most existing robotic navigation policies are trained on a single modality, limiting their adaptability to…

2026

Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Control

RSS 2026poster

Pretrained vision-language models (VLMs) can make semantic and visual inferences across diverse settings, providing valuable common-sense priors for robotic control. However, effectively grounding this knowledge in robot behaviors remains an open challenge. Prior methods often employ a hierarchical …

Cited by 0SourceScholar
2026

π∗0.6π0.6∗\pi^{*}_{0.6}: a VLA That Learns From Experience

RSS 2026poster

Vision–language–action (VLA) models offer a promising path toward general-purpose robots, but achieving the reliability and speed required for practical deployment remains challenging. We present a general-purpose method, RL with Experience and Corrections via Advantage-conditioned Policies (RECAP) …

Cited by 0SourceScholar
2024

LeLaN: Learning A Language-Conditioned Navigation Policy from In-the-Wild Video

CoRL 2024poster

We present our method, LeLaN, which uses action-free egocentric data to learn robust language-conditioned object navigation. By leveraging the knowledge of large vision and language models and grounding this knowledge using pre-trained segmentation and depth estimation models, we can label in-the-wi…

Cited by 7SourceScholar
2024

NoMaD: Goal Masked Diffusion Policies for Navigation and Exploration

ICRA 2024poster

Robotic learning for navigation in unfamiliar environments needs to provide policies for both task-oriented navigation (i.e., reaching a goal that the robot has located), and task-agnostic exploration (i.e., searching for a goal in a novel setting). Typically, these roles are handled by separate mod…

Cited by 124SourcecodeScholar
2024

Pushing the Limits of Cross-Embodiment Learning for Manipulation and Navigation

RSS 2024poster

Recent years in robotics and imitation learning have shown remarkable progress in training large-scale foundation models by leveraging data across a multitude of embodiments. The success of such policies might lead us to wonder: just how diverse can the robots in the training set be while still faci…