← Search

Zihan Xiong

3 accepted papers

2026

HiFICL: High-Fidelity In-Context Learning for Multimodal Tasks

CVPR 2026

In-Context Learning (ICL) is a significant paradigm for Large Multimodal Models (LMMs), using a few in-context demonstrations (ICDs) for new task adaptation. However, its performance is sensitive to demonstration configurations and computationally expensive. Mathematically, the influence of these de

Cited by 0SourcecodeScholar
2026

SupCLAP: Controlling Optimization Trajectory Drift in Audio-Text Contrastive Learning with Support Vector Regularization

ICLR 2026poster

Contrastive language-audio pretraining, which aims to unify multimodal representations in a shared embedding space, serves as a cornerstone for building a wide range of applications, from cross-modal retrieval to cutting-edge multimodal large language models. However, we find that the perpendicular…

Cited by 0SourceScholar