← Search

Kaixuan Wu

4 accepted papers

2026

TMR-VLA: Vision-Language-Action Model for Magnetic Motion Control of Tri-Leg Silicone-Based Soft Robot

ICRA 2026poster

In-vivo environments, magnetically actuated soft robots offer advantages such as wireless operation and precise control, showing promising potential for painless detection and therapeutic procedures. We developed a trileg magnetically driven soft robot (TMR) whose multi-legged design enables more fl…

Cited by 0Scholar
2025

AVQACL: A Novel Benchmark for Audio-Visual Question Answering Continual Learning

CVPR 2025poster

In this paper, a novel benchmark for audio-visual question answering continual learning (AVQACL) is introduced, aiming to study fine-grained scene understanding and spatial-temporal reasoning in videos under a continual learning setting. To facilitate this multimodal continual leaning task, we creat…

2024

Interpretable Short Video Rumor Detection Based on Modality Tampering

COLING 2024main

With the rapid development of social media and short video applications in recent years, browsing short videos has become the norm. Due to its large user base and unique appeal, spreading rumors via short videos has become a severe social problem. Many methods simply fuse multimodal features for rum…