← Search

Duojun Huang

3 accepted papers

2026

VCapsBench: A Large-scale Fine-grained Benchmark for Video Caption Quality Evaluation

AAAI 2026technical

Video captions play a crucial role in text-to-video generation tasks, as their quality directly influences the semantic coherence and visual fidelity of the generated videos. Although large vision-language models (VLMs) have demonstrated significant potential in caption generation, existing benchmar

Cited by 0SourcePDFScholar
2024

AlignSAM: Aligning Segment Anything Model to Open Context via Reinforcement Learning

CVPR 2024poster

Powered by massive curated training data Segment Anything Model (SAM) has demonstrated its impressive generalization capabilities in open-world scenarios with the guidance of prompts. However the vanilla SAM is class-agnostic and heavily relies on user-provided prompts to segment objects of interest…

2023

Divide and Adapt: Active Domain Adaptation via Customized Learning

CVPR 2023highlight

Active domain adaptation (ADA) aims to improve the model adaptation performance by incorporating the active learning (AL) techniques to label a maximally-informative subset of target samples. Conventional AL methods do not consider the existence of domain shift, and hence, fail to identify the truly…