← Search

Luojie Yang

2 accepted papers

2026

Learning a Unified Latent Action Space from Videos with Action-centric Cycle Consistency

CVPR 2026

Video data provides a rich source beyond expensive action-labeled data for advancing robot learning. Recent approaches have demonstrated promising potential in leveraging video data by learning latent actions for policy training. The latent action tokenizer encodes latent actions between successive

Cited by 0SourceScholar
2025

Open-RGBT: Open-Vocabulary RGB-T Zero-Shot Semantic Segmentation in Open-World Environments

ICRA 2025

Semantic segmentation is a critical technique for effective scene understanding. Traditional RGB-T semantic segmentation models often struggle to generalize across diverse scenarios due to their reliance on pretrained models and predefined categories. Recent advancements in Visual Language Models (V

Cited by 0SourcecodeScholar