← Search

Yinqiang Zheng

69 accepted papers

2026

HarmoQ: Harmonized Post-Training Quantization for High-Fidelity Image Super-Resolution

AAAI 2026technical

Post-training quantization offers an efficient pathway to deploy super-resolution models, yet existing methods treat weight and activation quantization independently, missing their critical interplay. Through controlled experiments on SwinIR, we uncover a striking asymmetry: weight quantization prim

Cited by 0SourcePDFScholar
2026

Motion-Aware Animatable Gaussian Avatars Deblurring

CVPR 2026

The creation of 3D human avatars from multi-view videos is a significant yet challenging task in computer vision. However, existing techniques rely on high-quality, sharp images as input, which are often impractical to obtain in real-world scenarios due to variations in human motion speed and intens

Cited by 0SourcecodeScholar
2025

Adversarial Attacks on Event-Based Pedestrian Detectors: A Physical Approach

AAAI 2025technical

Event cameras, known for their low latency and high dynamic range, show great potential in pedestrian detection applications. However, while recent research has primarily focused on improving detection accuracy, the robustness of event-based visual models against physical adversarial attacks has rec…

Cited by 0SourcePDFScholar
2025

Dr. RAW: Towards General High-Level Vision from RAW with Efficient Task Conditioning

NeurIPS 2025poster

We introduce Dr. RAW, a unified and tuning-efficient framework for high-level computer vision tasks directly operating on camera RAW data. Unlike previous approaches that optimize image signal processing (ISP) pipelines and fully fine-tune networks for each task, Dr. RAW achieves state-of-the-art pe…

Cited by 0SourceScholar
2025

Instruction-based Image Manipulation by Watching How Things Move

CVPR 2025highlight

This paper introduces a novel dataset construction pipeline that samples pairs of frames from videos and uses multimodal large language models (MLLMs) to generate editing instructions for training instruction-based image manipulation models. Video frames inherently preserve the identity of subjects…

Cited by 3SourcePDFScholar
2025

MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation

CVPR 2025poster

The image-to-video (I2V) generation is conditioned on the static image, which has been enhanced recently by the motion intensity as an additional control signal. These motion-aware models are appealing to generate diverse motion patterns, yet there lacks a reliable motion estimator for training such…

Cited by 4SourcePDFScholar
2025

Not All Degradations Are Equal: A Targeted Feature Denoising Framework for Generalizable Image Super-Resolution

ICCV 2025poster

Generalizable Image Super-Resolution aims to enhance model generalization capabilities under unknown degradations. To achieve such goal, the models are expected to focus only on image content-related features instead of degradation details (i.e., overfitting degradations).Recently, numerous approach…

Cited by 0SourcePDFScholar
2025

Random Is All You Need: Random Noise Injection on Feature Statistics for Generalizable Deep Image Denoising

ICLR 2025poster

Recent advancements in generalizable deep image denoising have catalyzed the development of robust noise-handling models. The current state-of-the-art, Masked Training (MT), constructs a masked swinir model which is trained exclusively on Gaussian noise ($\sigma$=15) but can achieve commendable deno…

Cited by 0SourcePDFScholar
2025

ResMaster: Mastering High-Resolution Image Generation via Structural and Fine-Grained Guidance

AAAI 2025technical

Diffusion models excel at producing high-quality images; however, scaling to higher resolutions, such as 4K, often results in structural distortions, and repetitive patterns. To this end, we introduce ResMaster, a novel, training-free method that empowers resolution-limited diffusion models to gener…

Cited by 11SourcePDFScholar
2025

Revolutionizing EMCCD Denoising through a Novel Physics-Based Learning Framework for Noise Modeling

ICLR 2025poster

Electron-multiplying charge-coupled device (EMCCD) has been instrumental in sensitive observations under low-light situations including astronomy, material science, and biology. Despite its ingenious designs to enhance target signals overcoming read-out circuit noises, produced images are not compl…

Cited by 0SourcePDFScholar
2025

SUICA: Learning Super-high Dimensional Sparse Implicit Neural Representations for Spatial Transcriptomics

ICML 2025poster

Spatial Transcriptomics (ST) is a method that captures gene expression profiles aligned with spatial coordinates. The discrete spatial distribution and the super-high dimensional sequencing results make ST data challenging to be modeled effectively. In this paper, we manage to model ST in a continuo…

2025

Towards Explicit Exoskeleton for the Reconstruction of Complicated 3D Human Avatars

ICCV 2025poster

In this paper, we highlight a critical yet often overlooked factor in most 3D human tasks, namely modeling complicated 3D human with with hand-held objects or loose-fitting clothing. It is known that the parameterized formulation of SMPL is able to fit human skin; while hand-held objects and loose-f…

2025

Tree-NeRV: Efficient Non-Uniform Sampling for Neural Video Representation via Tree-Structured Feature Grids

ICCV 2025poster

Implicit Neural Representations for Videos (NeRV) have emerged as a powerful paradigm for video representation, enabling direct mappings from frame indices to video frames. However, existing NeRV-based methods do not fully exploit temporal redundancy, as they rely on uniform sampling along the tempo…

2025

You Always Recognize Me (YARM): Robust Texture Synthesis Against Multi-View Corruption

ICML 2025poster

Damage to imaging systems and complex external environments often introduce corruption, which can impair the performance of deep learning models pretrained on high-quality image data. Previous methods have focused on restoring degraded images or fine-tuning models to adapt to out-of-distribution dat…

2024

Fooling Polarization-Based Vision using Locally Controllable Polarizing Projection

CVPR 2024poster

Polarization is a fundamental property of light that encodes abundant information regarding surface shape material illumination and viewing geometry. The computer vision community has witnessed a blossom of polarization-based vision applications such as reflection removal shape-from-polarization (Sf…

Cited by 1SourcePDFScholar
2024

IQ-VFI: Implicit Quadratic Motion Estimation for Video Frame Interpolation

CVPR 2024poster

Advanced video frame interpolation (VFI) algorithms approximate intermediate motions between two input frames to synthesize intermediate frame. However they struggle to handle complex scenarios with curvilinear motions since they overlook the latent acceleration information between the input frames.…

Cited by 7SourcePDFScholar
2024

Navigating Beyond Dropout: An Intriguing Solution towards Generalizable Image Super Resolution

CVPR 2024poster

Deep learning has led to a dramatic leap on Single Image Super-Resolution (SISR) performances in recent years. While most existing work assumes a simple and fixed degradation model (e.g. bicubic downsampling) the research of Blind SR seeks to improve model generalization ability with unknown degrada…

Cited by 4SourcePDFScholar
2024

Rolling Shutter Correction with Intermediate Distortion Flow Estimation

CVPR 2024poster

This paper proposes to correct the rolling shutter (RS) distorted images by estimating the distortion flow from the global shutter (GS) to RS directly. Existing methods usually perform correction using the undistortion flow from the RS to GS. They initially predict the flow from consecutive RS frame…

2024

Within the Dynamic Context: Inertia-aware 3D Human Modeling with Pose Sequence

ECCV 2024poster

"Neural rendering techniques have significantly advanced 3D human body modeling. However, previous approaches overlook dynamics induced by factors such as motion inertia, leading to challenges in scenarios where the pose remains static while the appearance changes, such as abrupt stops after spinnin…

Cited by 7SourcePDFScholar
2023

Assessor360: Multi-sequence Network for Blind Omnidirectional Image Quality Assessment

NeurIPS 2023poster

Blind Omnidirectional Image Quality Assessment (BOIQA) aims to objectively assess the human perceptual quality of omnidirectional images (ODIs) without relying on pristine-quality image information. It is becoming more significant with the increasing advancement of virtual reality (VR) technology. H…

2023

Blur Interpolation Transformer for Real-World Motion From Blur

CVPR 2023poster

This paper studies the challenging problem of recovering motion from blur, also known as joint deblurring and interpolation or blur temporal super-resolution. The challenges are twofold: 1) the current methods still leave considerable room for improvement in terms of visual quality even on the synth…

2023

HOTCOLD Block: Fooling Thermal Infrared Detectors with a Novel Wearable Design

AAAI 2023technical

Adversarial attacks on thermal infrared imaging expose the risk of related applications. Estimating the security of these systems is essential for safely deploying them in the real world. In many cases, realizing the attacks in the physical space requires elaborate special perturbations. These solut…

2023

High-Fidelity Event-Radiance Recovery via Transient Event Frequency

CVPR 2023poster

High-fidelity radiance recovery plays a crucial role in scene information reconstruction and understanding. Conventional cameras suffer from limited sensitivity in dynamic range, bit depth, and spectral response, etc. In this paper, we propose to use event cameras with bio-inspired silicon sensors,…

2023

MasaCtrl: Tuning-Free Mutual Self-Attention Control for Consistent Image Synthesis and Editing

ICCV 2023poster

Despite the success in large-scale text-to-image generation and text-conditioned image editing, existing methods still struggle to produce consistent generation and editing results. For example, generation approaches usually fail to synthesize multiple images of the same objects/characters but with…

Cited by 455PDFcodeScholar
2023

NeRFrac: Neural Radiance Fields through Refractive Surface

ICCV 2023poster

Neural Radiance Fields (NeRF) is a popular neural expression for novel view synthesis. By querying spatial points and view directions, a multilayer perceptron (MLP) can be trained to output the volume density and radiance at each point, which lets us render novel views of the scene. The original NeR…

Cited by 12PDFcodeScholar
2023

Rethinking Video Frame Interpolation from Shutter Mode Induced Degradation

ICCV 2023poster

Image restoration from various motion-related degradations, like blurry effects recorded by a global shutter (GS) and jello effects caused by a rolling shutter (RS), has been extensively studied. It has been recently recognized that such degradations encode temporal information, which can be exploit…

Cited by 8PDFScholar
2023

Visibility Constrained Wide-Band Illumination Spectrum Design for Seeing-in-the-Dark

CVPR 2023poster

Seeing-in-the-dark is one of the most important and challenging computer vision tasks due to its wide applications and extreme complexities of in-the-wild scenarios. Existing arts can be mainly divided into two threads: 1) RGB-dependent methods restore information using degraded RGB inputs only (e.g…

2022

Animation from Blur: Multi-modal Blur Decomposition with Motion Guidance

ECCV 2022poster

"We study the challenging problem of recovering detailed motion from a single motion-blurred image. Existing solutions to this problem estimate a single image sequence without considering the motion ambiguity for each region. Therefore, the results tend to converge to the mean of the multi-modal pos…

2022

Both Style and Fog Matter: Cumulative Domain Adaptation for Semantic Foggy Scene Understanding

CVPR 2022oral

Although considerable progress has been made in semantic scene understanding under clear weather, it is still a tough problem under adverse weather conditions, such as dense fog, due to the uncertainty caused by imperfect observations. Besides, difficulties in collecting and labeling foggy images hi…

Cited by 64PDFScholar
2022

Bringing Rolling Shutter Images Alive with Dual Reversed Distortion

ECCV 2022poster

"Rolling shutter (RS) distortion can be interpreted as the result of picking a row of pixels from instant global shutter (GS) frames over time during the exposure of the RS camera. This means that the information of each instant GS frame is partially, yet sequentially, embedded into the row-dependen…

2022

Efficient Video Deblurring Guided by Motion Magnitude

ECCV 2022poster

"Video deblurring is a highly under-constrained problem due to the spatially and temporally varying blur. An intuitive approach for video deblurring includes two steps: a) detecting the blurry region in the current frame; b) utilizing the information from clear regions in adjacent frames for current…

2022

Learning Adaptive Warping for Real-World Rolling Shutter Correction

CVPR 2022poster

This paper proposes a real-world rolling shutter (RS) correction dataset, BS-RSC, and a corresponding model to correct the RS frames in a distorted video. Mobile devices in the consumer market with CMOS-based sensors for video capture often result in rolling shutter effects when relative movements o…

Cited by 27PDFcodeScholar
2022

Neural Global Shutter: Learn To Restore Video From a Rolling Shutter Camera With Global Reset Feature

CVPR 2022poster

Most computer vision systems assume distortion-free images as inputs. The widely used rolling-shutter (RS) image sensors, however, suffer from geometric distortion when the camera and object undergo motion during capture. Extensive researches have been conducted on correcting RS distortions. However…

Cited by 15PDFcodeScholar
2021

4D Hyperspectral Photoacoustic Data Restoration With Reliability Analysis

CVPR 2021poster

Hyperspectral photoacoustic (HSPA) spectroscopy is an emerging bi-modal imaging technology that is able to show the wavelength-dependent absorption distribution of the interior of a 3D volume. However, HSPA devices have to scan an object exhaustively in the spatial and spectral domains; and the acqu…

Cited by 4PDFScholar
2021

Event-Based Bispectral Photometry Using Temporally Modulated Illumination

CVPR 2021poster

Analysis of bispectral difference plays a critical role in various applications that involve rays propagating in a light absorbing medium. In general, the bispectral difference is obtained by subtracting signals at two individual wavelengths captured by ordinary digital cameras, which tends to inher…

Cited by 15PDFScholar
2021

Learning To Reconstruct High Speed and High Dynamic Range Videos From Events

CVPR 2021poster

Event cameras are novel sensors that capture the dynamics of a scene asynchronously. Such cameras record event streams with much shorter response latency than images captured by conventional cameras, and are also highly sensitive to intensity change, which is brought by the triggering mechanism of e…

Cited by 69PDFScholar
2021

Multi-View 3D Reconstruction of a Texture-Less Smooth Surface of Unknown Generic Reflectance

CVPR 2021poster

Recovering the 3D geometry of a purely texture-less object with generally unknown surface reflectance (e.g. nonLambertian) is regarded as a challenging task in multiview reconstruction. The major obstacle revolves around establishing cross-view correspondences where photometric constancy is violated…

Cited by 31PDFcodeScholar
2021

Tuning IR-Cut Filter for Illumination-Aware Spectral Reconstruction From RGB

CVPR 2021poster

To reconstruct spectral signals from multi-channel observations, in particular trichromatic RGBs, has recently emerged as a promising alternative to traditional scanning-based spectral imager. It has been proven that the reconstruction accuracy relies heavily on the spectral response of the RGB came…

Cited by 14PDFScholar
2020

Beyond Intra-modality: A Survey of Heterogeneous Person Re-identification

IJCAI 2020poster

An efficient and effective person re-identification (ReID) system relieves the users from painful and boring video watching and accelerates the process of video analysis. Recently, with the explosive demands of practical applications, a lot of research efforts have been dedicated to heterogeneous pe…

2020

Efficient Spatio-Temporal Recurrent Neural Network for Video Deblurring

ECCV 2020poster

Real-time video deblurring still remains a challenging task due to the complexity of spatially and temporally varying blur itself and the requirement of low computational cost. To improve the network efficiency, we adopt residual dense blocks into RNN cells, so as to efficiently extract the spatial…

Cited by 175SourcePDFScholar
2020

Learn to Recover Visible Color for Video Surveillance in a Day

ECCV 2020poster

In silicon sensors, the interference between visible and near-infrared (NIR) signals is a crucial problem. For all-day video surveillance, commercial camera systems usually adopt auxiliary NIR cut filter and NIR LED illumination to selectively block or enhance NIR signal according to the surrounding…

2020

Optical Flow in the Dark

CVPR 2020poster

Many successful optical flow estimation methods have been proposed, but they become invalid when tested in dark scenes because low-light scenarios are not considered when they are designed and current optical flow benchmark datasets lack low-light samples. Even if we preprocess to enhance the dark i…

Cited by 69PDFScholar
2019

Hyperspectral Image Super-Resolution With Optimized RGB Guidance

CVPR 2019poster

To overcome the limitations of existing hyperspectral cameras on spatial/temporal resolution, fusing a low resolution hyperspectral image (HSI) with a high resolution RGB (or multispectral) image into a high resolution HSI has been prevalent. Previous methods for this fusion task usually emplo…

Cited by 108PDFcodeScholar
2019

Learning to Reduce Dual-Level Discrepancy for Infrared-Visible Person Re-Identification

CVPR 2019poster

Infrared-Visible person RE-IDentification (IV-REID) is a rising task. Compared to conventional person re-identification (re-ID), IV-REID concerns the additional modality discrepancy originated from the different imaging processes of spectrum cameras, in addition to the person's appearance discrepanc…

Cited by 522PDFcodeScholar
2018

Coded Illumination and Imaging for Fluorescence Based Classification

ECCV 2018poster

The quick detection of specific substances in objects such as produce items via non-destructive visual cues is vital to ensuring the quality and safety of consumer products. At the same time, it is well-known that the fluorescence excitation-emission characteristics of many organic objects can serve…

Cited by 4SourcePDFScholar
2018

Deeply Learned Filter Response Functions for Hyperspectral Reconstruction

CVPR 2018poster

Hyperspectral reconstruction from RGB imaging has recently achieved significant progress via sparse coding and deep learning. However, a largely ignored fact is that existing RGB cameras are tuned to mimic human richromatic perception, thus their spectral responses are not necessarily optimal for h…

Cited by 115SourcePDFScholar
2018

Joint Camera Spectral Sensitivity Selection and Hyperspectral Image Recovery

ECCV 2018poster

Hyperspectral image (HSI) recovery from a single RGB image has attracted much attention, whose performance has recently been shown to be sensitive to the camera spectral sensitivity (CSS). In this paper, we present an efficient convolutional neural network (CNN) based method, which can jointly selec…

Cited by 70SourcePDFScholar
2018

Simultaneous 3D Reconstruction for Water Surface and Underwater Scene

ECCV 2018poster

This paper presents the first approach for simultaneously recovering the 3D shape of both the wavy water surface and the moving underwater scene. A portable camera array system is constructed, which captures the scene from multiple viewpoints above the water. The correspondences across these cameras…

Cited by 43SourcePDFScholar
2018

Stereo relative pose from line and point feature triplets

ECCV 2018poster

Stereo relative pose problem lies at the core of stereo visual odometry systems that are used in many applications. In this work we present two minimal solvers for stereo relative pose. We specifically con- sider the case when a minimal set consist of three point or line features and each of them ha…

2017

A Microfacet-Based Reflectance Model for Photometric Stereo With Highly Specular Surfaces

ICCV 2017poster

A precise, stable and invertible model for surface reflectance is the key to the success of photometric stereo with real world materials. Recent developments in the field have enabled shape recovery techniques for surfaces of various types, but an effective solution to directly estimating the surfac…

Cited by 28PDFScholar
2017

From RGB to Spectrum for Natural Scenes via Manifold-Based Mapping

ICCV 2017poster

Spectral analysis of natural scenes can provide much more detailed information about the scene than an ordinary RGB camera. The richer information provided by hyperspectral images has been beneficial to numerous applications, such as understanding natural environmental changes and classifying plants…

Cited by 125PDFScholar
2016

Exploiting Spectral-Spatial Correlation for Coded Hyperspectral Image Restoration

CVPR 2016poster

Conventional scanning and multiplexing techniques for hyperspectral imaging suffer from limited temporal and/or spatial resolution. To resolve this issue, coding techniques are becoming increasingly popular in developing snapshot systems for high-resolution hyperspectral imaging. For such systems,…

Cited by 121PDFScholar
2015

Illumination and Reflectance Spectra Separation of a Hyperspectral Image Meets Low-Rank Matrix Factorization

CVPR 2015poster

This paper addresses the illumination and reflectance spectra separation (IRSS) problem of a hyperspectral image captured under general spectral illumination. The huge amount of pixels in a hypersepctral image poses tremendous challenges on computational efficiency, yet in turn offers greater color…

Cited by 46SourcePDFScholar
2015

Separating Fluorescent and Reflective Components by Using a Single Hyperspectral Image

ICCV 2015poster

This paper introduces a novel method to separate fluorescent and reflective components in the spectral domain. In contrast to existing methods, which require to capture two or more images under varying illuminations, we aim to achieve this separation task by using a single hyperspectral image. After…

Cited by 12PDFScholar