← Search

Karan Aggarwal

2 accepted papers

2024

Efficient Continual Pre-training for Building Domain Specific Large Language Models

ACL 2024findings

Large language models (LLMs) have demonstrated remarkable open-domain capabilities. LLMs tailored for a domain are typically trained entirely on domain corpus to excel at handling domain-specific tasks. In this work, we explore an alternative strategy of continual pre-training as a means to develop…

2023

ECG-QALM: Entity-Controlled Synthetic Text Generation using Contextual Q&A for NER

ACL 2023findings

Named Entity Recognition (NER) state-of-the-art methods requires high-quality labeled datasets. Issues such as scarcity of labeled data, under-representation of entities, and privacy concerns with using sensitive data for training, can be significant barriers. Generating synthetic data to train mode…

Cited by 2SourcePDFScholar