← Search

Guodong Ding

6 accepted papers

2026

LightAVSeg: Lightweight Audio-Visual Segmentation

ICML 2026poster

Audio-Visual Segmentation (AVS) targets pixel level localization of sounding emitting objects in videos. However, existing models rely on dense cross-modal attention with quadratic computational cost, limiting their suitability for resource efficient deployment. Most efficiency oriented methods focu…

Cited by 0SourceScholar
2026

MEDFACT-R1: TOWARDS FACTUAL MEDICAL REASONING VIA PSEUDO-LABEL AUGMENTATION

ICASSP 2026poster

Ensuring factual consistency and reliable reasoning remains a critical challenge for medical vision-language models. We introduce MEDFACT-R1, a two-stage framework that integrates external knowledge grounding with reinforcement learning to improve the factual medical reasoning. The first stage uses…

Cited by 0SourcePDFScholar
2022

Leveraging Action Affinity and Continuity for Semi-Supervised Temporal Action Segmentation

ECCV 2022poster

"We present a semi-supervised learning approach to the temporal action segmentation task. The goal of the task is to temporally detect and segment actions in long, untrimmed procedural videos, where only a small set of videos are densely labelled, and a large collection of videos are unlabelled. To…

Cited by 19SourcePDFScholar