← Search

Alkesh Patel

3 accepted papers

2023

Referring to Screen Texts with Voice Assistants

ACL 2023industry

Voice assistants help users make phone calls, send messages, create events, navigate and do a lot more. However assistants have limited capacity to understand their users’ context. In this work, we aim to take a step in this direction. Our work dives into a new experience for users to refer to phone…

Cited by 2SourcePDFScholar
2021

Generating Natural Questions from Images for Multimodal Assistants

ICASSP 2021accepted

Generating natural, diverse, and meaningful questions from images is an essential task for multimodal assistants as it confirms whether they have understood the object and scene in the images properly. The research in visual question answering (VQA) and visual question generation (VQG) is a great st…

Cited by 0SourceScholar
2021

Noise Robust Named Entity Understanding for Voice Assistants

NAACL 2021industry

Named Entity Recognition (NER) and Entity Linking (EL) play an essential role in voice assistant interaction, but are challenging due to the special difficulties associated with spoken user queries. In this paper, we propose a novel architecture that jointly solves the NER and EL tasks by combining…

Cited by 5SourcePDFScholar