← Search

Vaibhava Goel

10 accepted papers

2021

CNNBiF: CNN-based Bigram Features for Named Entity Recognition

EMNLP 2021finding

Transformer models fine-tuned with a sequence labeling objective have become the dominant choice for named entity recognition tasks. However, a self-attention mechanism with unconstrained length can fail to fully capture local dependencies, particularly when training data is limited. In this paper,…

2017

Self-Critical Sequence Training for Image Captioning

CVPR 2017oral

Recently it has been shown that policy-gradient methods for reinforcement learning can be utilized to train deep end-to-end systems directly on non-differentiable metrics for the task at hand. In this paper we consider the problem of optimizing image captioning systems using reinforcement learning,…

Cited by 2619PDFScholar
2015

Annealed dropout trained maxout networks for improved LVCSR

ICASSP 2015accepted

A significant barrier to progress in automatic speech recognition (ASR) capability is the empirical reality that techniques rarely “scale”-the yield of many apparently fruitful techniques rapidly diminishes to zero as the training criterion or decoder is strengthened, or the size of the training set…

Cited by 7SourceScholar
2015

Data augmentation for deep convolutional neural network acoustic modeling

ICASSP 2015accepted

This paper investigates data augmentation based on label-preserving transformations for deep convolutional neural network (CNN) acoustic modeling to deal with limited training data. We show how stochastic feature mapping (SFM) can be carried out when training CNN models with log-Mel features as inpu…

Cited by 0SourceScholar
2015

Evaluating Deep Scattering Spectra with deep neural networks on large scale spontaneous speech task

ICASSP 2015accepted

Deep Scattering Network features introduced for image processing have recently proved useful in speech recognition as an alternative to log-mel features for Deep Neural Network (DNN) acoustic models. Scattering features use wavelet decomposition directly producing log-frequency spectrograms which ar…

Cited by 0SourceScholar