← Search

Samuel Kessler

7 accepted papers

2026

Learning GUI Grounding with Spatial Reasoning from Visual Feedback

ICML 2026poster

Graphical User Interface (GUI) grounding is commonly framed as a coordinate prediction task – given a natural language instruction, generate on-screen coordinates for actions such as clicks and keystrokes. However, recent Vision Language Models (VLMs) often fail to predict accurate numeric coordinat…

Cited by 0SourceScholar
2024

Fisher Flow Matching for Generative Modeling over Discrete Data

NeurIPS 2024poster

Generative modeling over discrete data has recently seen numerous success stories, with applications spanning language modeling, biological sequence design, and graph-structured molecular data. The predominant generative modeling paradigm for discrete data is still autoregressive, with more recent a…

Cited by 15SourcePDFScholar
2022

An Adapter Based Pre-Training for Efficient and Scalable Self-Supervised Speech Representation Learning

ICASSP 2022accepted

We present a method for transferring pre-trained self-supervised (SSL) speech representations to multiple languages. There is an abundance of unannotated speech, so creating self-supervised representations from raw audio and fine-tuning on small annotated datasets is a promising direction to build s…

Cited by 0SourceScholar
2022

Efficient Adapter Transfer of Self-Supervised Speech Models for Automatic Speech Recognition

ICASSP 2022accepted

Self-supervised learning (SSL) is a powerful tool that allows learning of underlying representations from unlabeled data. Transformer based models such as wav2vec 2.0 and HuBERT are leading the field in the speech domain. Generally these models are fine-tuned on a small amount of labeled data for a…

Cited by 0SourceScholar
2022

Same State, Different Task: Continual Reinforcement Learning without Interference

AAAI 2022technical

Continual Learning (CL) considers the problem of training an agent sequentially on a set of tasks while seeking to retain performance on all previous tasks. A key challenge in CL is catastrophic forgetting, which arises when performance on a previously mastered task is reduced when learning a new ta…

2021

Hierarchical Indian buffet neural networks for Bayesian continual learning

UAI 2021poster

We place an Indian Buffet process (IBP) prior over the structure of a Bayesian Neural Network (BNN), thus allowing the complexity of the BNN to increase and decrease automatically. We further extend this model such that the prior on the structure of each hidden layer is shared globally across all la…

Cited by 30SourcePDFScholar