← Search

Luca M. Schulze Buschoff

5 accepted papers

2026

Can vision language models learn intuitive physics from interaction?

ICML 2026poster

Pre-trained vision language models do not have good intuitions about the physical world. Recent work has shown that supervised fine-tuning can improve model performance on simple physical tasks. However, fine-tuned models do not appear to learn robust physical rules that can generalize to new contex…

Cited by 0SourceScholar
2025

Testing the Limits of Fine-Tuning for Improving Visual Cognition in Vision Language Models

ICML 2025poster

Pre-trained vision language models still fall short of human visual cognition. In an effort to improve visual cognition and align models with human behavior, we introduce visual stimuli and human judgments on visual cognition tasks, allowing us to systematically evaluate performance across cognitive…

Cited by 0SourcePDFScholar
2025

metabench - A Sparse Benchmark of Reasoning and Knowledge in Large Language Models

ICLR 2025poster

Large Language Models (LLMs) vary in their abilities on a range of tasks. Initiatives such as the Open LLM Leaderboard aim to quantify these differences with several large benchmarks (sets of test items to which an LLM can respond either correctly or incorrectly). However, high correlations withi…

Cited by 0SourcePDFScholar
2023

The Acquisition of Physical Knowledge in Generative Neural Networks

ICML 2023poster

As children grow older, they develop an intuitive understanding of the physical processes around them. Their physical understanding develops in stages, moving along developmental trajectories which have been mapped out extensively in previous empirical research. Here, we investigate how the learning…

2022

Trivial or Impossible --- dichotomous data difficulty masks model differences (on ImageNet and beyond)

ICLR 2022poster

"The power of a generalization system follows directly from its biases" (Mitchell 1980). Today, CNNs are incredibly powerful generalisation systems---but to what degree have we understood how their inductive bias influences model decisions? We here attempt to disentangle the various aspects that det…