← Search

Yu-Yun Tseng

2 accepted papers

2025

Acknowledging Focus Ambiguity in Visual Questions

ICCV 2025poster

No published work on visual question answering (VQA) accounts for ambiguity regarding where the content described in the question is located in the image. To fill this gap, we introduce VQ-FocusAmbiguity, the first VQA dataset that visually grounds each plausible image region a question could refer…

Cited by 0SourcePDFScholar
2022

VizWiz-FewShot: Locating Objects in Images Taken by People with Visual Impairments

ECCV 2022poster

"We introduce a few-shot localization dataset originating from photographers who authentically were trying to learn about the visual content in the images they took. It includes over 8,000 segmentations of 100 categories in over 4,000 images that were taken by people with visual impairments. Compare…

Cited by 15SourcePDFScholar