← Search

Congxuan Zhang

6 accepted papers

2026

FlowAnyTime: Efficient Fine-tuning with Intra-Inter Frame Distillation for All-Weather Optical Flow Estimation

AAAI 2026technical

Motion estimation in degraded scenes has long been a significant challenge, primarily attributed to substantial scene variations and insufficient training data. Existing approaches typically address this limitation by incorporating additional training strategies or modifying network architectures wi

Cited by 0SourcePDFScholar
2025

ICIMG-Net: Inject Context Information to Motion Generation for Optical Flow Estimation

ICASSP 2025accepted

Although the overall performance of existing optical flow estimation methods has improved rapidly, motion discontinuities caused by large displacements and occlusions remain significant challenges for accurate optical flow estimation. To address this issue, we propose a novel Inject Context Informat…

Cited by 0SourceScholar
2025

Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations

EMNLP 2025

Large Vision-Language Models (LVLMs) suffer from serious hallucination problems, where the model-generated responses are inconsistent with the visual inputs. Existing hallucination mitigation methods are mainly based on preference alignment and require external human annotations or auxiliary models

2025

MotionFlow: Joint Motion Priors and Appearance Enhancement for High-Accuracy Optical Flow Estimation

ICASSP 2025accepted

Although optical flow estimation has improved significantly in recent years, large displacements and occlusions remain challenging for current methods due to motion discontinuities that may hinder accurate feature correspondences in these regions, leading to degraded performance. To address this cha…

Cited by 0SourceScholar
2022

Attention-Aware Learning for Hyperparameter Prediction in Image Processing Pipelines

ECCV 2022poster

"Between the imaging sensor and the image applications, the hardware image signal processing (ISP) pipelines reconstruct an RGB image from the sensor signal and feed it into downstream tasks. The processing blocks in ISPs depend on a set of tunable hyperparameters that have a complex interaction wit…

Cited by 13SourcePDFScholar
2022

Open-Vocabulary One-Stage Detection With Hierarchical Visual-Language Knowledge Distillation

CVPR 2022poster

Open-vocabulary object detection aims to detect novel object categories beyond the training set. The advanced open-vocabulary two-stage detectors employ instance-level visual-to-visual knowledge distillation to align the visual space of the detector with the semantic space of the Pre-trained Visual-…

Cited by 106PDFcodeScholar