← Search

Fabien Baradel

6 accepted papers

2024

Cross-view and Cross-pose Completion for 3D Human Understanding

CVPR 2024poster

Human perception and understanding is a major domain of computer vision which like many other vision subdomains recently stands to gain from the use of large models pre-trained on large datasets. We hypothesize that the most common pre-training strategy of relying on general purpose object-centric i…

Cited by 5SourcePDFScholar
2022

Filtered-CoPhy: Unsupervised Learning of Counterfactual Physics in Pixel Space

ICLR 2022oral

Learning causal relationships in high-dimensional data (images, videos) is a hard task, as they are often defined on low dimensional manifolds and must be extracted from complex signals dominated by appearance, lighting, textures and also spurious correlations in the data. We present a method for le…

Cited by 13SourcePDFScholar
2022

PoseGPT: Quantization-Based 3D Human Motion Generation and Forecasting

ECCV 2022poster

"We address the problem of action-conditioned generation of human motion sequences. Existing work falls into two categories: forecast models conditioned on observed past motions, or generative models conditioned action labels and duration only. In contrast, we generate motion conditioned on observat…

2020

CoPhy: Counterfactual Learning of Physical Dynamics

ICLR 2020spotlight

Understanding causes and effects in mechanical systems is an essential component of reasoning in the physical world. This work poses a new problem of counterfactual learning of object mechanics from visual input. We develop the CoPhy benchmark to assess the capacity of the state-of-the-art models f…

Cited by 113SourceScholar
2018

Glimpse Clouds: Human Activity Recognition From Unstructured Feature Points

CVPR 2018poster

We propose a method for human activity recognition from RGB data that does not rely on any pose information during test time, and does not explicitly calculate pose information internally. Instead, a visual attention module learns to predict glimpse sequences in each frame. These glimpses correspond…

2018

Object Level Visual Reasoning in Videos

ECCV 2018poster

Human activity recognition is typically addressed by training models to detect key concepts like global and local motion, features related to object classes present in the scene, as well as features related to the global context. The next open challenges in activity recognition require a level of un…