← Search

Maozheng Zhao

2 accepted papers

2026

FrameOracle: Learning What to See and How Much to See in Videos

ICML 2026poster

Vision-language models (VLMs) advance video understanding but operate under tight computational budgets, making performance dependent on selecting a small, high-quality subset of frames. Existing frame sampling strategies, such as uniform or fixed-budget selection, fail to adapt to variations in con…

Cited by 3SourceScholar
2017

Shadow Detection With Conditional Generative Adversarial Networks

ICCV 2017oral

We introduce scGAN, a novel extension of conditional Generative Adversarial Networks (GAN) tailored for the challenging problem of shadow detection in images. Previous methods for shadow detection focus on learning the local appearance of shadow regions, while using limited local context reasoning i…

Cited by 245PDFScholar