NeurIPS 2015poster21 citations

On the Accuracy of Self-Normalized Log-Linear Models

Jacob Andreas, Maxim Rabinovich, Michael I Jordan, Dan Klein

Abstract

Calculation of the log-normalizer is a major computational obstacle in applications of log-linear models with large output spaces. The problem of fast normalizer computation has therefore attracted significant attention in the theoretical and applied machine learning literature. In this paper, we analyze a recently proposed technique known as ``self-normalization'', which introduces a regularization term in training to penalize log normalizers for deviating from zero. This makes it possible to use unnormalized model scores as approximate probabilities. Empirical evidence suggests that self-normalization is extremely effective, but a theoretical understanding of why it should work, and how generally it can be applied, is largely lacking.We prove upper bounds on the loss in accuracy due to self-normalization, describe classes of input distributionsthat self-normalize easily, and construct explicit examples of high-variance input distributions. Our theoretical results make predictions about the difficulty of fitting self-normalized models to several classes of distributions, and we conclude with empirical validation of these predictions on both real and synthetic datasets.

BibTeX
@inproceedings{NIPS2015_4a08142c,
 author = {Andreas, Jacob and Rabinovich, Maxim and Jordan, Michael I and Klein, Dan},
 booktitle = {Advances in Neural Information Processing Systems},
 editor = {C. Cortes and N. Lawrence and D. Lee and M. Sugiyama and R. Garnett},
 pages = {},
 publisher = {Curran Associates, Inc.},
 title = {On the Accuracy of Self-Normalized Log-Linear Models},
 url = {https://proceedings.neurips.cc/paper_files/paper/2015/file/4a08142c38dbe374195d41c04562d9f8-Paper.pdf},
 volume = {28},
 year = {2015}
}
On the Accuracy of Self-Normalized Log-Linear Models · NeurIPS 2015