← Search

Vicente Ordonez*

1 accepted papers

2024

Grounding Language Models for Visual Entity Recognition

ECCV 2024poster

"We introduce , an Autoregressive model for Visual Entity Recognition. Our model extends an autoregressive Multimodal Large Language Model by employing retrieval augmented constrained generation. It mitigates low performance on out-of-domain entities while excelling in queries that require visual re…