← Search

Jianhua Wu

4 accepted papers

2026

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models

ICML 2026oral

Vision-Language Models (VLMs) perform well on commonsense reasoning tasks but struggle with visual spatial reasoning. Most existing solutions introduce extra 3D priors or external spatial encoders, which increase complexity and degrade the underlying VLMs' general-purpose capabilities after spatial …

Cited by 0SourceScholar
2025

Online Motion Generation via Tangential Sampling-Based MPC Around Nonconvex Obstacles

RA-L 2025

Collision-free motion planning has a well-established research history, but the majority of studies have been centered around Euclidean space and conducted offline. The primary challenge in online motion planning lies in circumventing local minima, which become more pronounced in configuration space

Cited by 0SourceScholar