← Search

Baochen Xiong

3 accepted papers

2025

Pilot: Building the Federated Multimodal Instruction Tuning Framework

AAAI 2025technical

In this paper, we explore a novel federated multimodal instruction tuning task(FedMIT), which is significant for collaboratively fine-tuning MLLMs on different types of multimodal instruction data on distributed devices. To solve the new task, we propose a federated multimodal instruction tuning fra…

Cited by 1SourcePDFScholar
2024

Modality-Collaborative Test-Time Adaptation for Action Recognition

CVPR 2024poster

Video-based Unsupervised Domain Adaptation (VUDA) method improves the generalization of the video model enabling it to be applied to action recognition tasks in different environments. However these methods require continuous access to source data during the adaptation process which are impractical…

Cited by 6SourcePDFScholar
2022

Cross-Modal Federated Human Activity Recognition via Modality-Agnostic and Modality-Specific Representation Learning

AAAI 2022technical

In this paper, we propose a new task of cross-modal federated human activity recognition (CMF-HAR), which is conducive to promote the large-scale use of the HAR model on more local devices. To address the new task, we propose a feature-disentangled activity recognition network (FDARN), which has fiv…

Cited by 30SourcePDFScholar