← Search

Roberto Henschel

6 accepted papers

2025

StreamingT2V: Consistent, Dynamic, and Extendable Long Video Generation from Text

CVPR 2025poster

Text-to-video diffusion models enable the generation of high-quality videos that follow text instructions, simplifying the process of producing diverse and individual content. Current methods excel in generating short videos (up to 16s), but produce hard-cuts when naively extended to long video synt…

2023

Text2Video-Zero: Text-to-Image Diffusion Models are Zero-Shot Video Generators

ICCV 2023oral

Recent text-to-video generation approaches rely on computationally heavy training and require large-scale video datasets. In this paper, we introduce a new task, zero-shot text-to-video generation, and propose a low-cost approach (without any training or optimization) by leveraging the power of exis…

Cited by 578PDFcodeScholar
2022

LMGP: Lifted Multicut Meets Geometry Projections for Multi-Camera Multi-Object Tracking

CVPR 2022poster

Multi-Camera Multi-Object Tracking is currently drawing attention in the computer vision field due to its superior performance in real-world applications such as video surveillance with crowded scenes or in wide spaces. In this work, we propose a mathematically elegant multi-camera multiple object t…

Cited by 44PDFcodeScholar
2021

Making Higher Order MOT Scalable: An Efficient Approximate Solver for Lifted Disjoint Paths

ICCV 2021poster

We present an efficient approximate message passing solver for the lifted disjoint paths problem (LDP), a natural but NP-hard model for multiple object tracking (MOT). Our tracker scales to very large instances that come from long and crowded MOT sequences. Our approximate solver enables us to proce…

Cited by 45PDFcodeScholar
2020

Lifted Disjoint Paths with Application in Multiple Object Tracking

ICML 2020poster

We present an extension to the disjoint paths problem in which additional lifted edges are introduced to provide path connectivity priors. We call the resulting optimization problem the lifted disjoint paths problem. We show that this problem is NP-hard by reduction from integer multicommodity flow…

2018

Recovering Accurate 3D Human Pose in The Wild Using IMUs and a Moving Camera

ECCV 2018poster

In this work, we propose a method that combines a single hand-held camera and a set of Inertial Measurement Units (IMUs) attached at the body limbs to estimate accurate 3D poses in the wild. This poses many new challenges: the moving camera, heading drift, cluttered background, occlusions and many p…

Cited by 1257SourcePDFScholar