← Search

Yandu Chen

2 accepted papers

2026

IntentionVLA: Generalizable and Efficient Embodied Intention Reasoning for Human–Robot Interaction

ICRA 2026poster

Vision-Language-Action (VLA) models leverage pretrained vision-language models (VLMs) to couple perception with robotic control, offering a promising path toward general purpose embodied intelligence. However, current SOTA VLAs are primarily pretrained on multimodal tasks with limited relevance to e…

2025

Riemann-based Multi-scale Attention Reasoning Network for Text-3D Retrieval

AAAI 2025technical

Due to the challenges in acquiring paired Text-3D data and the inherent irregularity of 3D data structures, combined representation learning of 3D point clouds and text remains unexplored. In this paper, we propose a novel Riemann-based Multi-scale Attention Reasoning Network (RMARN) for text-3D ret…