← Search

Hung-Kai Chung

2 accepted papers

2026

SEASON: Mitigating Temporal Hallucination in Video Large Language Models via Self-Diagnostic Contrastive Decoding

CVPR 2026

Video Large Language Models (VideoLLMs) have shown remarkable progress in video understanding. However, these models still struggle to effectively perceive and exploit rich temporal information in videos when responding to user queries. Therefore, they often generate descriptions of events that are

Cited by 0SourcecodeScholar
2025

VideoMage: Multi-Subject and Motion Customization of Text-to-Video Diffusion Models

CVPR 2025poster

Customized text-to-video generation aims to produce high-quality videos that incorporate user-specified subject identities or motion patterns. However, existing methods mainly focus on personalizing a single concept, either subject identity or motion pattern, limiting their effectiveness for multipl…

Cited by 0SourcePDFScholar