← Search

Niranjan Uma Naresh

3 accepted papers

2022

Adversarial Data Augmentation for Task-Specific Knowledge Distillation of Pre-trained Transformers

AAAI 2022technical

Deep and large pre-trained language models (e.g., BERT, GPT-3) are state-of-the-art for various natural language processing tasks. However, the huge size of these models brings challenges to fine-tuning and online deployment due to latency and cost constraints. Existing knowledge distillation method…

Cited by 16SourcePDFScholar
2022

PENTATRON: PErsonalized coNText-Aware Transformer for Retrieval-based cOnversational uNderstanding

EMNLP 2022industry

Conversational understanding is an integral part of modern intelligent devices. In a large fraction of the global traffic from customers using smart digital assistants, frictions in dialogues may be attributed to incorrect understanding of the entities in a customer’s query due to factors including…

Cited by 6SourcePDFScholar
2019

Guaranteed Scalable Learning of Latent Tree Models

UAI 2019poster

We present an integrated approach to structure and parameter estimation in latent tree graphical models, where some nodes are hidden. Our overall approach follows a “divide-and-conquer” strategy that learns models over small groups of variables and iteratively merges into a global solution. The s…

Cited by 10SourcePDFScholar