← Search

Zhihao Shi

9 accepted papers

2026

CamDirector: Towards Long-Term Coherent Video Trajectory Editing

CVPR 2026

Video (camera) trajectory editing aims to synthesize new videos that follow user-defined camera paths while preserving scene content and plausibly inpainting previously unseen regions, upgrading amateur footage into professionally styled videos. Existing VTE methods struggle with precise camera cont

Cited by 0SourceScholar
2025

Accurate and Scalable Graph Neural Networks via Message Invariance

ICLR 2025poster

Message passing-based graph neural networks (GNNs) have achieved great success in many real-world applications. For a sampled mini-batch of target nodes, the message passing process is divided into two parts: message passing between nodes within the batch (MP-IB) and message passing from nodes outsi…

2025

IMFine: 3D Inpainting via Geometry-guided Multi-view Refinement

CVPR 2025poster

Current 3D inpainting and object removal methods are largely limited to front-facing scenes, facing substantial challenges when applied to diverse, "unconstrained" scenes where the camera orientation and trajectory are unrestricted. To bridge this gap, we introduce a novel approach that produces inp…

Cited by 0SourcePDFScholar
2025

Robust Neural Rendering in the Wild with Asymmetric Dual 3D Gaussian Splatting

NeurIPS 2025spotlight

3D reconstruction from in-the-wild images remains a challenging task due to inconsistent lighting conditions and transient distractors. Existing methods typically rely on heuristic strategies to handle the low-quality training data, which often struggle to produce stable and consistent reconstructio…

Cited by 0SourceScholar
2025

Wasserstein Style Distribution Analysis and Transform for Stylized Image Generation

ICCV 2025poster

Large-scale text-to-image diffusion models have achieved remarkable success in image generation, thereby driving the development of stylized image generation technologies. Recent studies introduce style information by empirically replacing specific features in attention blocks with style features. H…

Cited by 0SourcePDFScholar
2024

TexGen: Text-Guided 3D Texture Generation with Multi-view Sampling and Resampling

ECCV 2024poster

"Given a 3D mesh, we aim to synthesize 3D textures that correspond to arbitrary textual descriptions. Current methods for generating and assembling textures from sampled views often result in prominent seams or excessive smoothing. To tackle these issues, we present TexGen, a novel multi-view sampli…

2023

LMC: Fast Training of GNNs via Subgraph Sampling with Provable Convergence

ICLR 2023top-25%

The message passing-based graph neural networks (GNNs) have achieved great success in many real-world applications. However, training GNNs on large-scale graphs suffers from the well-known neighbor explosion problem, i.e., the exponentially increasing dependencies of nodes with the number of message…

2019

GridDehazeNet: Attention-Based Multi-Scale Network for Image Dehazing

ICCV 2019poster

We propose an end-to-end trainable Convolutional Neural Network (CNN), named GridDehazeNet, for single image dehazing. The GridDehazeNet consists of three modules: pre-processing, backbone, and post-processing. The trainable pre-processing module can generate learned inputs with better diversity and…

Cited by 1096PDFcodeScholar