← Search

Steven J. Rennie

3 accepted papers

2017

Self-Critical Sequence Training for Image Captioning

CVPR 2017oral

Recently it has been shown that policy-gradient methods for reinforcement learning can be utilized to train deep end-to-end systems directly on non-differentiable metrics for the task at hand. In this paper we consider the problem of optimizing image captioning systems using reinforcement learning,…

Cited by 2619PDFScholar
2015

Annealed dropout trained maxout networks for improved LVCSR

ICASSP 2015accepted

A significant barrier to progress in automatic speech recognition (ASR) capability is the empirical reality that techniques rarely “scale”-the yield of many apparently fruitful techniques rapidly diminishes to zero as the training criterion or decoder is strengthened, or the size of the training set…

Cited by 7SourceScholar