← Search

Michael Ogezi

2 accepted papers

2024

Semantically-Prompted Language Models Improve Visual Descriptions

NAACL 2024findings

Language-vision models like CLIP have made significant strides in vision tasks, such as zero-shot image classification (ZSIC). However, generating specific and expressive visual descriptions remains challenging; descriptions produced by current methods are often ambiguous and lacking in granularity.…

Cited by 0SourcePDFScholar