← Search

Matko Bošnjak

6 accepted papers

2024

Improving fine-grained understanding in image-text pre-training

ICML 2024poster

We introduce SPARse fine-grained Contrastive alignment (SPARC), a simple method for pretraining more fine-grained multimodal representations from image-text pairs. Given that multiple image patches often correspond to single words, we propose to learn a grouping of image patches for every token in t…

Cited by 16SourcePDFScholar
2023

SemPPL: Predicting Pseudo-Labels for Better Contrastive Representations

ICLR 2023poster

Learning from large amounts of unsupervised data and a small amount of supervision is an important open problem in computer vision. We propose a new semi-supervised learning method, Semantic Positives via Pseudo-Labels (SEMPPL), that combines labelled and unlabelled data to learn informative represe…

2022

Making Sense of Raw Input (Extended Abstract)

IJCAI 2022poster

How should a machine intelligence perform unsupervised structure discovery over streams of sensory input? One approach to this problem is to cast it as an apperception task. Here, the task is to construct an explicit interpretable theory that both explains the sensory sequence and also satisfies a s…

Cited by 0SourcePDFScholar
2018

SCAN: Learning Hierarchical Compositional Visual Concepts

ICLR 2018poster

The seemingly infinite diversity of the natural world arises from a relatively small set of coherent rules, such as the laws of physics or chemistry. We conjecture that these rules give rise to regularities that can be discovered through primarily unsupervised experiences and represented as abstract…

Cited by 151SourcePDFScholar
2017

Programming With a Differentiable Forth Interpreter

ICLR 2017workshop

There are families of neural networks that can learn to compute any function, provided sufficient training data. However, given that in practice training data is scarce for all but a small set of problems, a core question is how to incorporate prior knowledge into a model. Here we consider the case…

Cited by 121SourceScholar
2017

Programming with a Differentiable Forth Interpreter

ICML 2017poster

Given that in practice training data is scarce for all but a small set of problems, a core question is how to incorporate prior knowledge into a model. In this paper, we consider the case of prior procedural knowledge for neural networks, such as knowing how a program should traverse a sequence, but…

Cited by 121SourcePDFScholar