← Search

Jiazi Wang

1 accepted papers

2026

VaseVQA-3D: Benchmarking 3D VLMs on Ancient Greek Pottery

ICLR 2026poster

Vision-Language Models (VLMs) have achieved significant progress in multimodal understanding tasks, demonstrating strong capabilities particularly in general tasks such as image captioning and visual reasoning. However, when dealing with specialized cultural heritage domains like 3D vase artifacts,…

Cited by 0SourcecodeScholar