← Search

Goker Erdogan

3 accepted papers

2025

LayerLock: Non-collapsing Representation Learning with Progressive Freezing

ICCV 2025poster

We introduce LayerLock, a simple yet effective approach for self-supervised visual representation learning, that gradually transitions throughout training from predicting shallow features to deeper ones through progressive layer freezing. First, we make the observation that during training of video…

Cited by 0SourcePDFScholar
2024

Improving fine-grained understanding in image-text pre-training

ICML 2024poster

We introduce SPARse fine-grained Contrastive alignment (SPARC), a simple method for pretraining more fine-grained multimodal representations from image-text pairs. Given that multiple image patches often correspond to single words, we propose to learn a grouping of image patches for every token in t…

Cited by 16SourcePDFScholar
2021

SIMONe: View-Invariant, Temporally-Abstracted Object Representations via Unsupervised Video Decomposition

NeurIPS 2021poster

To help agents reason about scenes in terms of their building blocks, we wish to extract the compositional structure of any given scene (in particular, the configuration and characteristics of objects comprising the scene). This problem is especially difficult when scene structure needs to be inferr…

Cited by 82SourcePDFScholar