AAAI 2026technical0 citations
MegaCoin: Enhancing Medium-Grained Color Perception for Vision-Language Models
Ming-Chang Chiu, Shicheng Wen, Pin-Yu Chen, Xuezhe Ma
Abstract
In vision-language models (VLMs), the ability to perceive and interpret color and physical environment is crucial for achieving contextually accurate understanding and interaction. However, despite advances in multimodal modeling, there remains a significant lack of specialized datasets that rigorously evaluate a model
BibTeX
@inproceedings{aaai2026_megacoinenhancin,
title = {MegaCoin: Enhancing Medium-Grained Color Perception for Vision-Language Models},
author = {Ming-Chang Chiu and Shicheng Wen and Pin-Yu Chen and Xuezhe Ma},
booktitle = {AAAI 2026},
year = {2026}
}