← Search

Thomas J. Cashman

7 accepted papers

2025

DAViD: Data-efficient and Accurate Vision Models from Synthetic Data

ICCV 2025poster

The state of the art in human-centric computer vision achieves high accuracy and robustness across a diverse range of tasks. The most effective models in this domain have billions of parameters, thus requiring extremely large datasets, expensive training regimes, and compute-intensive inference. In…

Cited by 0SourcePDFScholar
2025

VoluMe - Authentic 3D Video Calls from Live Gaussian Splat Prediction

ICCV 2025poster

Virtual 3D meetings offer the potential to enhance copresence, increase engagement and thus improve effectiveness of remote meetings compared to standard 2D video calls. However, representing people in 3D meetings remains a challenge; existing solutions achieve high quality by using complex hardware…

Cited by 4SourcePDFScholar
2022

FLAG: Flow-Based 3D Avatar Generation From Sparse Observations

CVPR 2022poster

To represent people in mixed reality applications for collaboration and communication, we need to generate realistic and faithful avatar poses. However, the signal streams that can be applied for this task from head-mounted devices (HMDs) are typically limited to head pose and hand pose estimates. W…

Cited by 59PDFScholar
2021

Fake It Till You Make It: Face Analysis in the Wild Using Synthetic Data Alone

ICCV 2021poster

We demonstrate that it is possible to perform face-related computer vision in the wild using synthetic data alone. The community has long enjoyed the benefits of synthesizing training data with graphics, but the domain gap between real and synthetic data has remained a problem, especially for human…

Cited by 339PDFcodeScholar
2021

Full-Body Motion From a Single Head-Mounted Device: Generating SMPL Poses From Partial Observations

ICCV 2021poster

The increased availability and maturity of head-mounted and wearable devices opens up opportunities for remote communication and collaboration. However, the signal streams provided by these devices (e.g., head pose, hand pose, and gaze direction) do not represent a whole person. One of the main open…

Cited by 69PDFScholar
2020

The Phong Surface: Efficient 3D Model Fitting using Lifted Optimization

ECCV 2020poster

Realtime perceptual and interaction capabilities in mixed reality require a range of 3D tracking problems to be solved at low latency on resource-constrained hardware such as head-mounted devices. Indeed, for devices such as HoloLens 2 where the CPU and GPU are left available for applications, multi…

Cited by 17SourcePDFScholar
2017

An Efficient Background Term for 3D Reconstruction and Tracking With Smooth Surface Models

CVPR 2017poster

We present a novel strategy to shrink and constrain a 3D model, represented as a smooth spline-like surface, within the visual hull of an object observed from one or multiple views. This new 'background' or 'silhouette' term combines the efficiency of previous approaches based on an image-plane dist…

Cited by 7PDFScholar