← Search

Miaomiao Cui

17 accepted papers

2025

GaussianIP: Identity-Preserving Realistic 3D Human Generation via Human-Centric Diffusion Prior

CVPR 2025poster

Text-guided 3D human generation has advanced with the development of efficient 3D representations and 2D-lifting methods like score distillation sampling (SDS). However, current methods suffer from prolonged training times and often produce results that lack fine facial and garment details. In this…

2025

MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling

CVPR 2025poster

Character video synthesis aims to produce realistic videos of animatable characters within lifelike scenes. As a fundamental problem in the computer vision and graphics community, 3D works typically require multi-view captures for per-case training, which severely limits their applicability of model…

Cited by 18SourcePDFScholar
2025

Perception-as-Control: Fine-grained Controllable Image Animation with 3D-aware Motion Representation

ICCV 2025poster

Motion-controllable image animation is a fundamental task with a wide range of potential applications. Recent works have made progress in controlling camera or object motion via various motion representations, while they still struggle to support collaborative camera and object motion control with a…

Cited by 0SourcePDFScholar
2025

Towards Fine-grained Interactive Segmentation in Images and Videos

ICCV 2025poster

The recent Segment Anything Models (SAMs) have emerged as foundational visual models for general interactive segmentation. Despite demonstrating robust generalization abilities, they still suffer from performance degradations in scenarios that demand accurate masks. Existing methods for high-precisi…

Cited by 0SourcePDFScholar
2024

3DToonify: Creating Your High-Fidelity 3D Stylized Avatar Easily from 2D Portrait Images

CVPR 2024poster

Visual content creation has aroused a surge of interest given its applications in mobile photography and AR/VR. Portrait style transfer and 3D recovery from monocular images as two representative tasks have so far evolved independently. In this paper we make a connection between the two and tackle t…

Cited by 2SourcePDFScholar
2024

DiffusionGAN3D: Boosting Text-guided 3D Generation and Domain Adaptation by Combining 3D GANs and Diffusion Priors

CVPR 2024poster

Text-guided domain adaptation and generation of 3D-aware portraits find many applications in various fields. However due to the lack of training data and the challenges in handling the high variety of geometry and appearance the existing methods for these tasks suffer from issues like inflexibility…

2024

En3D: An Enhanced Generative Model for Sculpting 3D Humans from 2D Synthetic Data

CVPR 2024poster

We present En3D an enhanced generative scheme for sculpting high-quality 3D human avatars. Unlike previous works that rely on scarce 3D datasets or limited 2D collections with imbalanced viewing angles and imprecise pose priors our approach aims to develop a zero-shot 3D generative scheme capable of…

Cited by 10SourcePDFScholar
2024

InitNO: Boosting Text-to-Image Diffusion Models via Initial Noise Optimization

CVPR 2024poster

Recent strides in the development of diffusion models exemplified by advancements such as Stable Diffusion have underscored their remarkable prowess in generating visually compelling images. However the imperative of achieving a seamless alignment between the generated image and the provided prompt…

2023

A Hierarchical Representation Network for Accurate and Detailed Face Reconstruction From In-the-Wild Images

CVPR 2023poster

Limited by the nature of the low-dimensional representational capacity of 3DMM, most of the 3DMM-based face reconstruction (FR) methods fail to recover high-frequency facial details, such as wrinkles, dimples, etc. Some attempt to solve the problem by introducing detail maps or non-linear operations…

2022

ABPN: Adaptive Blend Pyramid Network for Real-Time Local Retouching of Ultra High-Resolution Photo

CVPR 2022poster

Photo retouching finds many applications in various fields. However, most existing methods are designed for global retouching and seldom pay attention to the local region, while the latter is actually much more tedious and time-consuming in photography pipelines. In this paper, we propose a novel ad…

Cited by 14PDFcodeScholar
2022

Active Boundary Loss for Semantic Segmentation

AAAI 2022technical

This paper proposes a novel active boundary loss for semantic segmentation. It can progressively encourage the alignment between predicted boundaries and ground-truth boundaries during end-to-end training, which is not explicitly enforced in commonly used cross-entropy loss. Based on the predicted b…

2022

Box-Supervised Instance Segmentation with Level Set Evolution

ECCV 2022poster

"In contrast to the fully supervised methods using pixel-wise mask labels, box-supervised instance segmentation takes advantage of the simple box annotations, which has recently attracted a lot of research attentions. In this paper, we propose a novel single-shot box-supervised instance segmentation…

2022

Structure-Aware Flow Generation for Human Body Reshaping

CVPR 2022poster

Body reshaping is an important procedure in portrait photo retouching. Due to the complicated structure and multifarious appearance of human bodies, existing methods either fall back on the 3D domain via body morphable model or resort to keypoint-based image deformation, leading to inefficiency and…

Cited by 7PDFcodeScholar
2022

Unpaired Cartoon Image Synthesis via Gated Cycle Mapping

CVPR 2022poster

In this paper, we present a general-purpose solution to cartoon image synthesis with unpaired training data. In contrast to previous works learning pre-defined cartoon styles for specified usage scenarios (portrait or scene), we aim to train a common cartoon translator which can not only simultaneou…

Cited by 21PDFScholar
2021

PPR10K: A Large-Scale Portrait Photo Retouching Dataset With Human-Region Mask and Group-Level Consistency

CVPR 2021poster

Different from general photo retouching tasks, portrait photo retouching (PPR), which aims to enhance the visual quality of a collection of flat-looking portrait photos, has its special and practical requirements such as human-region priority (HRP) and group-level consistency (GLC). HRP requires tha…

Cited by 57PDFcodeScholar
2020

Boosting Semantic Human Matting With Coarse Annotations

CVPR 2020oral

Semantic human matting aims to estimate the per-pixel opacity of the foreground human regions. It is quite challenging that usually requires user interactive trimaps and plenty of high quality annotated data. Annotating such kind of data is labor intensive and requires great skills beyond normal use…

Cited by 113PDFScholar