← Search

Jean-Charles Bazin

15 accepted papers

2026

Large-scale Codec Avatars: The Unreasonable Effectiveness of Large-scale Avatar Pretraining

CVPR 2026

High-quality 3D avatar modeling faces a critical trade-off between fidelity and generalization. On the one hand, multi-view studio data enables high-fidelity modeling of humans with precise control over expressions and poses, but it struggles to generalize to real-world data due to limited scale and

Cited by 0SourcecodeScholar
2024

Doubly Hierarchical Geometric Representations for Strand-based Human Hairstyle Generation

NeurIPS 2024poster

We introduce a doubly hierarchical generative representation for strand-based 3D hairstyle geometry that progresses from coarse, low-pass filtered guide hair to densely populated hair strands rich in high-frequency details. We employ the Discrete Cosine Transform (DCT) to separate low-frequency stru…

Cited by 0SourcePDFScholar
2024

FAMOUS: High-Fidelity Monocular 3D Human Digitization Using View Synthesis

ECCV 2024poster

"The advancement in deep implicit modeling and articulated models has significantly enhanced the process of digitizing human figures in 3D from just a single image. While state-of-the-art methods have greatly improved geometric precision, the challenge of accurately inferring texture remains, partic…

2023

MMG-Ego4D: Multimodal Generalization in Egocentric Action Recognition

CVPR 2023poster

In this paper, we study a novel problem in egocentric action recognition, which we term as "Multimodal Generalization" (MMG). MMG aims to study how systems can generalize when data from certain modalities is limited or even completely missing. We thoroughly investigate MMG in the context of standard…

2023

PATMAT: Person Aware Tuning of Mask-Aware Transformer for Face Inpainting

ICCV 2023poster

Generative models such as StyleGAN2 and Stable Diffusion have achieved state-of-the-art performance in computer vision tasks such as image synthesis, inpainting, and de-noising. However, current generative models for face inpainting often fail to preserve fine facial details and the identity of the…

Cited by 3PDFcodeScholar
2020

Robust and Efficient Estimation of Absolute Camera Pose for Monocular Visual Odometry

ICRA 2020poster

Given a set of 3D-to-2D point correspondences corrupted by outliers, we aim to robustly estimate the absolute camera pose. Existing methods robust to outliers either fail to guarantee high robustness and efficiency simultaneously, or require an appropriate initial pose and thus lack generality. In c…

Cited by 5SourceScholar
2019

Leveraging Structural Regularity of Atlanta World for Monocular SLAM

ICRA 2019poster

A wide range of man-made environments can be abstracted as the Atlanta world. It consists of a set of Atlanta frames with a common vertical (gravitational) axis and multiple horizontal axes orthogonal to this vertical axis. This paper focuses on leveraging the regularity of Atlanta world for monocul…

Cited by 48SourceScholar
2019

Line-based Absolute and Relative Camera Pose Estimation in Structured Environments

IROS 2019poster

3D lines in structured environments encode particular regularity like parallelism and orthogonality. We leverage this structural regularity to estimate the absolute and relative camera poses. We decouple the rotation and translation, and propose a novel rotation estimation method. We decompose the a…

Cited by 22SourceScholar
2019

Quasi-Globally Optimal and Efficient Vanishing Point Estimation in Manhattan World

ICCV 2019oral

The image lines projected from parallel 3D lines intersect at a common point called the vanishing point (VP). Manhattan world holds for the scenes with three orthogonal VPs. In Manhattan world, given several lines in a calibrated image, we aim at clustering them by three unknown-but-sought VPs. The…

Cited by 37PDFScholar
2018

A Monocular SLAM System Leveraging Structural Regularity in Manhattan World

ICRA 2018poster

The structural features in Manhattan world encode useful geometric information of parallelism, orthogonality and/or coplanarity in the scene. By fully exploiting these structural features, we propose a novel monocular SLAM system which provides accurate estimation of camera poses and 3D map. The for…

Cited by 72SourceScholar
2018

Globally Optimal Inlier Set Maximization for Atlanta Frame Estimation

CVPR 2018poster

In this work, we describe man-made structures via an appropriate structure assumption, called Atlanta world, which contains a vertical direction (typically the gravity direction) and a set of horizontal directions orthogonal to the vertical direction. Contrary to the commonly used Manhattan world as…

Cited by 22SourcePDFScholar
2018

Robust Camera Pose Estimation via Consensus on Ray Bundle and Vector Field

IROS 2018poster

Estimating the camera pose requires point correspondences. However, in practice, correspondences are inevitably corrupted by outliers, which affects the pose estimation. We propose a general and accurate outlier removal strategy for robust camera pose estimation. The proposed strategy can detect out…

Cited by 7SourceScholar
2015

FaceDirector: Continuous Control of Facial Performance in Video

ICCV 2015poster

We present a method to continuously blend between multiple facial performances of an actor, which can contain different facial expressions or emotional states. As an example, given sad and angry video takes of a scene, our method empowers the movie director to specify arbitrary weighted combinations…

Cited by 19PDFScholar