← Search

Zixuan Ni

2 accepted papers

2026

Beyond Sequential Tools: A Unified VLM Agent System for Photographic Post-Processing via Dynamic Multi-Expert Fusion

CVPR 2026

Real-world image restoration is challenged by complex, coupled degradations. Existing "all-in-one" models often lack generalization, while agentic systems suffer from inefficient sequential tool invocation. We propose a VLM-guided one-shot framework for universal photographic post-processing. Our sy

Cited by 0SourcecodeScholar
2023

Continual Vision-Language Representation Learning with Off-Diagonal Information

ICML 2023poster

Large-scale multi-modal contrastive learning frameworks like CLIP typically require a large amount of image-text samples for training. However, these samples are always collected continuously in real scenarios. This paper discusses the feasibility of continual CLIP training using streaming data. Unl…

Cited by 23SourcePDFScholar