← Search

Zhihui Yang

1 accepted papers

2025

Does Visual Grounding Enhance the Understanding of Embodied Knowledge in Large Language Models?

EMNLP 2025

Despite significant progress in multimodal language models (LMs), it remains unclear whether visual grounding enhances their understanding of embodied knowledge compared to text-only models. To address this question, we propose a novel embodied knowledge understanding benchmark based on the perceptu

Cited by 0SourcePDFScholar