← Search

Borja Rodríguez Gálvez

4 accepted papers

2025

An Information-Theoretic Analysis of Thompson Sampling with Infinite Action Spaces

ICASSP 2025accepted

This paper studies the Bayesian regret of the Thompson Sampling algorithm for bandit problems, building on the information-theoretic framework introduced by Russo and Van Roy [1]. Specifically, it extends the rate-distortion analysis of Dong and Van Roy [2], which provides near-optimal bounds for li…

Cited by 0SourceScholar
2025

Information-Theoretic Minimax Regret Bounds for Reinforcement Learning based on Duality

ICASSP 2025accepted

We study agents acting in an unknown environment where the agent’s goal is to find a robust policy. We consider robust policies as policies that achieve high cumulative rewards for all possible environments. To this end, we consider agents minimizing the maximum regret over different environment par…

Cited by 0SourceScholar
2023

The Role of Entropy and Reconstruction in Multi-View Self-Supervised Learning

ICML 2023poster

The mechanisms behind the success of multi-view self-supervised learning (MVSSL) are not yet fully understood. Contrastive MVSSL methods have been studied through the lens of InfoNCE, a lower bound of the Mutual Information (MI). However, the relation between other MVSSL methods and MI remains uncle…

2021

Tighter Expected Generalization Error Bounds via Wasserstein Distance

NeurIPS 2021poster

This work presents several expected generalization error bounds based on the Wasserstein distance. More specifically, it introduces full-dataset, single-letter, and random-subset bounds, and their analogous in the randomized subsample setting from Steinke and Zakynthinou [1]. Moreover, when the loss…

Cited by 52SourcePDFScholar