← Search

Jimmy S. Ren

16 accepted papers

2025

DiT4SR: Taming Diffusion Transformer for Real-World Image Super-Resolution

ICCV 2025poster

Large-scale pre-trained diffusion models are becoming increasingly popular in solving the Real-World Image Super-Resolution (Real-ISR) problem because of their rich generative priors. The recent development of diffusion transformer (DiT) has witnessed overwhelming performance over the traditional UN…

Cited by 0SourcePDFScholar
2025

Event-guided HDR Reconstruction with Diffusion Priors

ICCV 2025poster

Events provide High Dynamic Range (HDR) intensity change which can guide Low Dynamic Range (LDR) image for HDR reconstruction. However, events only provide temporal intensity differences and it is still ill-posed in over-/under-exposed areas due to missing initial reference brightness and color info…

2024

Latency Correction for Event-guided Deblurring and Frame Interpolation

CVPR 2024poster

Event cameras with their high temporal resolution dynamic range and low power consumption are particularly good at time-sensitive applications like deblurring and frame interpolation. However their performance is hindered by latency variability especially under low-light conditions and with fast-mov…

Cited by 9SourcePDFScholar
2023

Range-Nullspace Video Frame Interpolation With Focalized Motion Estimation

CVPR 2023poster

Continuous-time video frame interpolation is a fundamental technique in computer vision for its flexibility in synthesizing motion trajectories and novel video frames at arbitrary intermediate time steps. Yet, how to infer accurate intermediate motion and synthesize high-quality video frames are two…

Cited by 6SourcePDFScholar
2022

Deep Bayesian Video Frame Interpolation

ECCV 2022poster

"We present deep Bayesian video frame interpolation, a novel approach for upsampling a low frame-rate video temporally to its higher frame-rate counterpart. Our approach learns posterior distributions of optical flows and frames to be interpolated, which is optimized via learned gradient descent for…

2021

Bringing Events Into Video Deblurring With Non-Consecutively Blurry Frames

ICCV 2021poster

Recently, video deblurring has attracted considerable research attention, and several works suggest that events at high time rate can benefit deblurring. In this paper, we develop a principled framework D2Nets for video deblurring to exploit non-consecutively blurry frames, and propose a flexible ev…

Cited by 73PDFcodeScholar
2021

Learning a Non-Blind Deblurring Network for Night Blurry Images

CVPR 2021poster

Deblurring night blurry images is difficult, because the common-used blur model based on the linear convolution operation does not hold in this situation due to the influence of saturated pixels. In this paper, we propose a non-blind deblurring network (NBDN) to restore night blurry images. To mitig…

Cited by 36PDFScholar
2021

Training Weakly Supervised Video Frame Interpolation With Events

ICCV 2021poster

Event-based video frame interpolation is promising as event cameras capture dense motion signals that can greatly facilitate motion-aware synthesis. However, training existing frameworks for this task requires high frame-rate videos with synchronized events, posing challenges to collect real trainin…

Cited by 42PDFcodeScholar
2020

EfficientFCN: Holistically-guided Decoding for Semantic Segmentation

ECCV 2020poster

Both performance and efficiency are important to semantic segmentation. State-of-the-art semantic segmentation algorithms are mostly based on dilated Fully Convolutional Networks (dilatedFCN), which adopt dilated convolutions in the backbone networks to extract high-resolution feature maps for achie…

Cited by 74SourcePDFScholar
2020

Learning to Predict Context-adaptive Convolution for Semantic Segmentation

ECCV 2020poster

Long-range contextual information is essential for achieving high-performance semantic segmentation. Previous feature re-weighting methods demonstrate that using global context for re-weighting feature channels can effectively improve the accuracy of semantic segmentation. However, the globally-shar…

Cited by 37SourcePDFScholar
2020

PIPAL: a Large-Scale Image Quality Assessment Dataset for Perceptual Image Restoration

ECCV 2020poster

Image quality assessment (IQA) is the key factor for the fast development of image restoration (IR) algorithms. The most recent IR methods based on Generative Adversarial Networks (GANs) have achieved significant improvement in visual performance, but also presented great challenges for quantitative…

Cited by 232SourcePDFScholar
2019

DAVANet: Stereo Deblurring With View Aggregation

CVPR 2019oral

Nowadays stereo cameras are more commonly adopted in emerging devices such as dual-lens smartphones and unmanned aerial vehicles. However, they also suffer from blurry images in dynamic scenes which leads to visual discomfort and hampers further image processing. Previous works have succeeded in mon…

Cited by 119PDFScholar
2019

Spatially Variant Linear Representation Models for Joint Filtering

CVPR 2019poster

Joint filtering mainly uses an additional guidance image as a prior and transfers its structures to the target image in the filtering process. Different from existing algorithms that rely on locally linear models or hand-designed objective functions to extract the structural information from the gui…

Cited by 53PDFScholar
2019

Structure-Preserving Stereoscopic View Synthesis With Multi-Scale Adversarial Correlation Matching

CVPR 2019poster

This paper addresses stereoscopic view synthesis from a single image. Various recent works solve this task by reorganizing pixels from the input view to reconstruct the target one in a stereo setup. However, purely depending on such photometric-based reconstruction process, the network may produce s…

Cited by 12PDFScholar