← Search

Zijing Zhao

7 accepted papers

2026

SIGMark: Scalable In-Generation Watermark with Blind Extraction for Video Diffusion

ICLR 2026poster

Artificial Intelligence Generated Content (AIGC), particularly video generation with diffusion models, has been advanced rapidly. Invisible watermarking is a key technology for protecting AI-generated videos and tracing harmful content, and thus plays a crucial role in AI safety. Beyond post-proces…

Cited by 0SourceScholar
2025

AR-VRM: Imitating Human Motions for Visual Robot Manipulation with Analogical Reasoning

ICCV 2025poster

Visual Robot Manipulation (VRM) aims to enable a robot to follow natural language instructions based on robot states and visual observations, and therefore requires costly multi- modal data. To compensate for the deficiency of robot data, existing approaches have employed vision-language pre- traini…

2025

PlanLLM: Video Procedure Planning with Refinable Large Language Models

AAAI 2025technical

Video procedure planning, i.e., planning a sequence of action steps given the video frames of start and goal states, is an essential ability for embodied AI. Recent works utilize Large Language Models (LLMs) to generate enriched action step description texts to guide action step decoding. Although L…

2025

Zero Shot Domain Adaptive Semantic Segmentation by Synthetic Data Generation and Progressive Adaptation

IROS 2025

Deep learning-based semantic segmentation models achieve impressive results yet remain limited in handling distribution shifts between training and test data. In this paper, we present SDGPA (Synthetic Data Generation and Progressive Adaptation), a novel method that tackles zero-shot domain adaptive

Cited by 3SourcecodeScholar
2023

Masked Retraining Teacher-Student Framework for Domain Adaptive Object Detection

ICCV 2023poster

Domain adaptive Object Detection (DAOD) leverages a labeled domain (source) to learn an object detector generalizing to a novel domain without annotation (target). Recent advances use a teacher-student framework, i.e., a student model is supervised by the pseudo labels from a teacher model. Though g…

Cited by 31PDFcodeScholar
2015

An Accurate Iris Segmentation Framework Under Relaxed Imaging Constraints Using Total Variation Model

ICCV 2015poster

This paper proposes a novel and more accurate iris segmentation framework to automatically segment iris region from the face images acquired with relaxed imaging under visible or near-infrared illumination, which provides strong feasibility for applications in surveillance, forensics and the search…

Cited by 168PDFScholar