← Search

Shafiq R. Joty

2 accepted papers

2021

Preventing Early Endpointing for Online Automatic Speech Recognition

ICASSP 2021accepted

With the recent development of end-to-end models in speech recognition, there have been more interests in adapting these models for online speech recognition. However, using end-to-end models for online speech recognition is known to suffer from an early endpointing problem, which brings in many del…

Cited by 0SourceScholar
2018

Look, Imagine and Match: Improving Textual-Visual Cross-Modal Retrieval With Generative Models

CVPR 2018poster

Textual-visual cross-modal retrieval has been a hot research topic in both computer vision and natural language processing communities. Learning appropriate representations for multi-modal data is crucial for the cross-modal retrieval performance. Unlike existing image-text retrieval approaches that…

Cited by 476SourcePDFScholar