AAAI 2026technical0 citations

MegaCoin: Enhancing Medium-Grained Color Perception for Vision-Language Models

Ming-Chang Chiu, Shicheng Wen, Pin-Yu Chen, Xuezhe Ma

Abstract

In vision-language models (VLMs), the ability to perceive and interpret color and physical environment is crucial for achieving contextually accurate understanding and interaction. However, despite advances in multimodal modeling, there remains a significant lack of specialized datasets that rigorously evaluate a model

BibTeX
@inproceedings{aaai2026_megacoinenhancin,
  title = {MegaCoin: Enhancing Medium-Grained Color Perception for Vision-Language Models},
  author = {Ming-Chang Chiu and Shicheng Wen and Pin-Yu Chen and Xuezhe Ma},
  booktitle = {AAAI 2026},
  year = {2026}
}