← Search

Lennart Rietdorf

1 accepted papers

2025

Does VLM Classification Benefit from LLM Description Semantics?

AAAI 2025technical

Accurately describing images with text is a foundation of explainable AI. Vision-Language Models (VLMs) like CLIP have recently addressed this by aligning images and texts in a shared embedding space, expressing semantic similarities between vision and language embeddings. VLM classification can be…