2026
VaseVQA-3D: Benchmarking 3D VLMs on Ancient Greek Pottery
ICLR 2026poster
Vision-Language Models (VLMs) have achieved significant progress in multimodal understanding tasks, demonstrating strong capabilities particularly in general tasks such as image captioning and visual reasoning. However, when dealing with specialized cultural heritage domains like 3D vase artifacts,…