← Search

Chongyan Chen

4 accepted papers

2025

Acknowledging Focus Ambiguity in Visual Questions

ICCV 2025poster

No published work on visual question answering (VQA) accounts for ambiguity regarding where the content described in the question is located in the image. To fill this gap, we introduce VQ-FocusAmbiguity, the first VQA dataset that visually grounds each plausible image region a question could refer…

Cited by 0SourcePDFScholar
2025

mmWalk: Towards Multi-modal Multi-view Walking Assistance

NeurIPS 2025poster

Walking assistance in extreme or complex environments remains a significant challenge for people with blindness or low vision (BLV), largely due to the lack of a holistic scene understanding. Motivated by the real-world needs of the BLV community, we build mmWalk, a simulated multi-modal dataset tha…

Cited by 0SourceScholar