← Search

Samuel Yu

3 accepted papers

2024

Language Models as Black-Box Optimizers for Vision-Language Models

CVPR 2024poster

Vision-language models (VLMs) pre-trained on web-scale datasets have demonstrated remarkable capabilities on downstream tasks when fine-tuned with minimal data. However many VLMs rely on proprietary data and are not open-source which restricts the use of white-box approaches for fine-tuning. As such…

2023

Multimodality Helps Unimodality: Cross-Modal Few-Shot Learning With Multimodal Models

CVPR 2023poster

The ability to quickly learn a new task with minimal instruction - known as few-shot learning - is a central aspect of intelligent agents. Classical few-shot benchmarks make use of few-shot samples from a single modality, but such samples may not be sufficient to characterize an entire concept class…

2022

PACS: A Dataset for Physical Audiovisual Commonsense Reasoning

ECCV 2022poster

"In order for AI to be safely deployed in real-world scenarios such as hospitals, schools, and the workplace, it must be able to robustly reason about the physical world. Fundamental to this reasoning is physical common sense: understanding the physical properties and affordances of available object…