← Search

Isinsu Katircioglu

4 accepted papers

2025

GEM: A Generalizable Ego-Vision Multimodal World Model for Fine-Grained Ego-Motion, Object Dynamics, and Scene Composition Control

CVPR 2025poster

We present GEM, a Generalizable Ego-vision Multimodal world model that predicts future frames using a reference frame, sparse features, human poses, and ego-trajectories. Hence, our model has precise control over object dynamics, ego-agent motion and human poses. GEM generates paired RGB and depth o…

2021

Human Detection and Segmentation via Multi-View Consensus

ICCV 2021poster

Self-supervised detection and segmentation of foreground objects aims for accuracy without annotated training data. However, existing approaches predominantly rely on restrictive assumptions on appearance and motion. For scenes with dynamic activities and camera motion, we propose a multi-camera fra…

Cited by 3PDFcodeScholar
2019

Neural Scene Decomposition for Multi-Person Motion Capture

CVPR 2019poster

Learning general image representations has proven key to the success of many computer vision tasks. For example, many approaches to image understanding problems rely on deep networks that were initially trained on ImageNet, mostly because the learned features are a valuable starting point to learn f…

Cited by 61PDFScholar
2018

Learning Monocular 3D Human Pose Estimation From Multi-View Images

CVPR 2018poster

Accurate 3D human pose estimation from single images is possible with sophisticated deep-net architectures that have been trained on very large datasets. However, this still leaves open the problem of capturing motions for which no such database exists. Manual annotation is tedious, slow, and error…

Cited by 308SourcePDFScholar