← Search

Jeff Tan

5 accepted papers

2026

Contact-guided Real2Sim from Monocular Video with Planar Scene Primitives

ICLR 2026poster

We introduce CRISP, a method that recovers simulatable human motion and scene geometry from monocular video. Prior work on joint human--scene reconstruction relies on data-driven priors and joint optimization with no physics in the loop, or recovers noisy geometry with artifacts that cause motion-tr…

Cited by 0SourcecodeScholar
2025

DiffusionSfM: Predicting Structure and Motion via Ray Origin and Endpoint Diffusion

CVPR 2025poster

Current Structure-from-Motion (SfM) methods typically follow a two-stage pipeline, combining learned or geometric pairwise reasoning with a subsequent global optimization step. In contrast, we propose a data-driven multi-view reasoning approach that directly infers 3D scene geometry and camera poses…

2025

MonoFusion: Sparse-View 4D Reconstruction via Monocular Fusion

ICCV 2025poster

We address the problem of dynamic scene reconstruction from sparse-view videos. Prior work often requires dense multi-view captures with hundreds of calibrated cameras (e.g. Panoptic Studio) - such multi-view setups are prohibitively expensive to build and cannot capture diverse scenes in-the-wild.…

Cited by 0SourcePDFScholar
2024

Upgrading Search Applications in the Era of LLMs: A Demonstration with Practical Lessons

IJCAI 2024poster

While traditional search systems have mostly been satisfactorily relying on lexical based sparse retrievers such as BM25, recent research advances in neural models, the current day large language models (LLMs) hold good promise for practical search applications as well. In this work, we discuss a co…

Cited by 0SourcePDFScholar