← Search

Minh Hieu Phan

3 accepted papers

2025

OVG-HQ: Online Video Grounding with Hybrid-modal Queries

ICCV 2025poster

Video grounding (VG) task focuses on locating specific moments in a video based on a query, usually in text form. However, traditional VG struggles with some scenarios like streaming video or queries using visual cues. To fill this gap, we present a new task named Online Video Grounding with Hybrid-…

Cited by 0SourcePDFScholar
2024

CARER - ClinicAl Reasoning-Enhanced Representation for Temporal Health Risk Prediction

EMNLP 2024main

The increasing availability of multimodal data from electronic health records (EHR) has paved the way for deep learning methods to improve diagnosis accuracy. However, deep learning models are data-driven, requiring large-scale datasets to achieve high generalizability. Inspired by how human experts…

2022

Class Similarity Weighted Knowledge Distillation for Continual Semantic Segmentation

CVPR 2022poster

Deep learning models are known to suffer from the problem of catastrophic forgetting when they incrementally learn new classes. Continual learning for semantic segmentation (CSS) is an emerging field in computer vision. We identify a problem in CSS: A model tends to be confused between old and new c…

Cited by 65PDFScholar