2024
Visual-Textual Entailment with Quantities Using Model Checking and Knowledge Injection
COLING 2024main
In recent years, there has been great interest in multimodal inference. We concentrate on visual-textual entailment (VTE), a critical task in multimodal inference. VTE is the task of determining entailment relations between an image and a sentence. Several deep learning-based approaches have been pr…