← Search

Gabriel Goh

4 accepted papers

2025

Scaling and evaluating sparse autoencoders

ICLR 2025oral

Sparse autoencoders provide a promising unsupervised approach for extracting interpretable features from a language model by reconstructing activations from a sparse bottleneck layer. Since language models learn many concepts, autoencoders need to be very large to recover all relevant features. Howe…

2021

Learning Transferable Visual Models From Natural Language Supervision

ICML 2021oral

State-of-the-art computer vision systems are trained to predict a fixed set of predetermined object categories. This restricted form of supervision limits their generality and usability since additional labeled data is needed to specify any other visual concept. Learning directly from raw text about…

2021

Zero-Shot Text-to-Image Generation

ICML 2021spotlight

Text-to-image generation has traditionally focused on finding better modeling assumptions for training on a fixed dataset. These assumptions might involve complex architectures, auxiliary losses, or side information such as object part labels or segmentation masks supplied during training. We descri…

2016

Satisfying Real-world Goals with Dataset Constraints

NeurIPS 2016poster

The goal of minimizing misclassification error on a training set is often just one of several real-world goals that might be defined on different datasets. For example, one may require a classifier to also make positive predictions at some specified rate for some subpopulation (fairness), or to achi…