← Search

Gizem Yüce

4 accepted papers

2025

Learning In-context $n$-grams with Transformers: Sub-$n$-grams Are Near-Stationary Points

ICML 2025poster

In this article, we explore the loss landscape of next-token prediction with transformers. Specifically, we focus on learning in-context n-gram language models with cross-entropy loss using a simplified two-layer transformer. We design a series of transformers that represent $k$-grams (for $k \leq n…

Cited by 0SourcePDFScholar
2025

Learning Parametric Distributions from Samples and Preferences

ICML 2025spotlight

Recent advances in language modeling have underscored the role of preference feedback in enhancing model performance. This paper investigates the conditions under which preference feedback improves parameter estimation in classes of continuous parametric distributions. In our framework, the learner…

2023

Can semi-supervised learning use all the data effectively? A lower bound perspective

NeurIPS 2023spotlight

Prior theoretical and empirical works have established that semi-supervised learning algorithms can leverage the unlabeled data to improve over the labeled sample complexity of supervised learning (SL) algorithms. However, existing theoretical work focuses on regimes where the unlabeled data is suff…

Cited by 0SourcePDFScholar
2022

A Structured Dictionary Perspective on Implicit Neural Representations

CVPR 2022poster

Implicit neural representations (INRs) have recently emerged as a promising alternative to classical discretized representations of signals. Nevertheless, despite their practical success, we still do not understand how INRs represent signals. We propose a novel unified perspective to theoretically a…

Cited by 95PDFcodeScholar