2026
Make LVLMs Focus: Context-Aware Attention Modulation for Better Multimodal In-Context Learning
AAAI 2026technical
Multimodal in-context learning (ICL) is becoming a key capability that allows large vision-language models (LVLMs) to adapt to novel tasks without parameter updates, which expands their usefulness in many real-world applications. However, ICL performance remains unstable even when the in-context dem