← Search

Ege Ozguroglu

4 accepted papers

2024

Dreamitate: Real-World Visuomotor Policy Learning via Video Generation

CoRL 2024poster

A key challenge in manipulation is learning a policy that can robustly generalize to diverse visual environments. A promising mechanism for learning robust policies is to leverage video generative models, which are pretrained on large-scale datasets of internet videos. In this paper, we propose a vi…

Cited by 26SourceScholar
2024

EraseDraw : Learning to Insert Objects by Erasing Them from Images

ECCV 2024poster

"Creative processes such as painting often involve creating different components of an image one by one. Can we build a computational model to perform this task? Prior works often fail by making global changes to the image, inserting objects in unrealistic spatial locations, and generating inaccurat…

Cited by 2SourcePDFScholar
2024

Generative Camera Dolly: Extreme Monocular Dynamic Novel View Synthesis

ECCV 2024oral

"Accurate reconstruction of complex dynamic scenes from just a single viewpoint continues to be a challenging task in computer vision. Current dynamic novel view synthesis methods typically require videos from many different camera viewpoints, necessitating careful recording setups, and significantl…

Cited by 23SourcePDFScholar
2024

pix2gestalt: Amodal Segmentation by Synthesizing Wholes

CVPR 2024highlight

We introduce pix2gestalt a framework for zero-shot amodal segmentation which learns to estimate the shape and appearance of whole objects that are only partially visible behind occlusions. By capitalizing on large-scale diffusion models and transferring their representations to this task we learn a…