← Search

Cennet Oguz

5 accepted papers

2026

Grounding or Guessing? Visual Signals for Detecting Hallucinations in Sign Language Translation

ICLR 2026poster

Hallucination, where models generate fluent text unsupported by visual evidence, remains a major flaw in vision–language models and is especially critical in sign language translation (SLT). In SLT, meaning depends on precise grounding in video, and gloss-free models are particularly vulnerable beca…

Cited by 1SourceScholar
2024

MMAR: Multilingual and Multimodal Anaphora Resolution in Instructional Videos

EMNLP 2024finding

Multilingual anaphora resolution identifies referring expressions and implicit arguments in texts and links to antecedents that cover several languages. In the most challenging setting, cross-lingual anaphora resolution, training data, and test data are in different languages. As knowledge needs to…

2023

Find-2-Find: Multitask Learning for Anaphora Resolution and Object Localization

EMNLP 2023long main

In multimodal understanding tasks, visual and linguistic ambiguities can arise. Visual ambiguity can occur when visual objects require a model to ground a referring expression in a video without strong supervision, while linguistic ambiguity can occur from changes in entities in action flows. As an…

Cited by 0SourceScholar
2023

InterroLang: Exploring NLP Models and Datasets through Dialogue-based Explanations

EMNLP 2023long findings

While recently developed NLP explainability methods let us open the black box in various ways (Madsen et al., 2022), a missing ingredient in this endeavor is an interactive tool offering a conversational interface. Such a dialogue system can help users explore datasets and models with explanations i…

Cited by 22SourcecodeScholar
2021

Synthesis of Compositional Animations From Textual Descriptions

ICCV 2021poster

How can we animate 3D-characters from a movie script or move robots by simply telling them what we would like them to do?" How unstructured and complex can we make a sentence and still generate plausible movements from it?" These are questions that need to be answered in the long-run, as the field i…

Cited by 202PDFcodeScholar