← Search

Guohua Gao

2 accepted papers

2026

Seeing Is Believing: Grounding Long-Video Understanding in Spatio-Temporal Visual Evidence

AAAI 2026technical

Although Vision Language Models (VLMs) have excelled at image and video understanding, applying them to hour-long videos is held back by two interrelated challenges: exorbitant computational expense and a qualitative breakdown in long-term temporal reasoning. Thus, models tend to generate answers ba

Cited by 0SourcePDFScholar
2019

Modal Dynamics and Analysis of a Vertical Stretch-Retractable Continuum Manipulator with Large Deflection

ICRA 2019poster

Efficient and reliable dynamic modelling and analysis is critical to the shape control and estimate of a backbone continuum manipulator. This paper presents a novel dynamic modelling method to investigate the deformation modal properties of a vertical stretch-retractable continuum manipulator (VSRCM…

Cited by 1SourceScholar