← Search

Haiyong Zheng

10 accepted papers

2025

Face Forgery Video Detection via Temporal Forgery Cue Unraveling

CVPR 2025poster

Face Forgery Video Detection (FFVD) is a critical yet challenging task in determining whether a digital facial video is authentic or forged. Existing FFVD methods typically focus on isolated spatial or coarsely fused spatiotemporal information, failing to leverage temporal forgery cues thus resultin…

2024

Spherical Pseudo-Cylindrical Representation for Omnidirectional Image Super-resolution

AAAI 2024technical

Omnidirectional images have attracted significant attention in recent years due to the rapid development of virtual reality technologies. Equirectangular projection (ERP), a naive form to store and transfer omnidirectional images, however, is challenging for existing two-dimensional (2D) image super…

Cited by 6SourcePDFScholar
2024

Video Harmonization with Triplet Spatio-Temporal Variation Patterns

CVPR 2024poster

Video harmonization is an important and challenging task that aims to obtain visually realistic composite videos by automatically adjusting the foreground's appearance to harmonize with the background. Inspired by the short-term and long-term gradual adjustment process of manual harmonization we pre…

2021

Image Harmonization With Transformer

ICCV 2021poster

Image harmonization, aiming to make composite images look more realistic, is an important and challenging task. The composite, synthesized by combining foreground from one image with background from another image, inevitably suffers from the issue of inharmonious appearance caused by distinct imagin…

Cited by 91PDFcodeScholar
2021

Multi-Modal Multi-Action Video Recognition

ICCV 2021poster

Multi-action video recognition is much more challenging due to the requirement to recognize multiple actions co-occurring simultaneously or sequentially. Modeling multi-action relations is beneficial and crucial to understand videos with multiple actions, and actions in a video are usually presented…

Cited by 12PDFcodeScholar
2020

CoTeRe-Net: Discovering Collaborative Ternary Relations in Videos

ECCV 2020poster

Modeling relations is crucial to understand videos for action and behavior recognition. Current relation models mainly reason about relations of invisibly implicit cues, while important relations of visually explicit cues are rarely considered, and the collaboration between them is usually ignored.…

2020

Spiral Generative Network for Image Extrapolation

ECCV 2020poster

In this paper, motivated by human natural ability to perceive unseen surroundings imaginatively, we propose a novel Spiral Generative Network, SpiralNet, to perform image extrapolation in a spiral manner, which regards extrapolation as an evolution process growing from an input sub-image along a spi…

2018

Discriminative Region Proposal Adversarial Networks for High-Quality Image-to-Image Translation

ECCV 2018poster

Image-to-image translation has been made much progress with embracing Generative Adversarial Networks (GANs). However, it's still very challenging for translation tasks that require high quality, especially at high-resolution and photorealism. In this paper, we present Discriminative Region Proposal…