← Search

Eduard Dragut

10 accepted papers

2026

DanceHA: A Multi-Agent Framework for Document-Level Aspect-Based Sentiment Analysis

AAAI 2026technical

Aspect-Based Sentiment Intensity Analysis (ABSIA) has garnered increasing attention, though research largely focuses on domain-specific, sentence-level settings. In contrast, document-level ABSIA--particularly in addressing complex tasks like extracting Aspect-Category-Opinion-Sentiment-Intensity (A

Cited by 0SourcePDFScholar
2025

DynClean: Training Dynamics-based Label Cleaning for Distantly-Supervised Named Entity Recognition

NAACL 2025findings

Distantly Supervised Named Entity Recognition (DS-NER) has attracted attention due to its scalability and ability to automatically generate labeled data. However, distant annotation introduces many mislabeled instances, limiting its performance. Most of the existing work attempt to solve this proble…

2025

Taxonomy-Driven Knowledge Graph Construction for Domain-Specific Scientific Applications

ACL 2025finding

We present a taxonomy-driven framework for constructing domain-specific knowledge graphs (KGs) that integrates structured taxonomies, Large Language Models (LLMs) and Retrieval-Augmented Generation (RAG). Although we focus on climate science to illustrate its effectiveness, our approach can potentia…

2025

UniT: One Document, Many Revisions, Too Many Edit Intention Taxonomies

ACL 2025finding

Writing is inherently iterative, each revision enhancing information representation. One revision may contain many edits. Examination of the intentions behind edits provides valuable insights into an editor’s expertise, the dynamics of collaborative writing, and the evolution of a document. Current…

2024

SciDMT: A Large-Scale Corpus for Detecting Scientific Mentions

COLING 2024main

We present SciDMT, an enhanced and expanded corpus for scientific mention detection, offering a significant advancement over existing related resources. SciDMT contains annotated scientific documents for datasets (D), methods (M), and tasks (T). The corpus consists of two components: 1) the SciDMT m…

2024

SciER: An Entity and Relation Extraction Dataset for Datasets, Methods, and Tasks in Scientific Documents

EMNLP 2024main

Scientific information extraction (SciIE) is critical for converting unstructured knowledge from scholarly articles into structured data (entities and relations). Several datasets have been proposed for training and validating SciIE models. However, due to the high complexity and cost of annotating…

2022

COIN – an Inexpensive and Strong Baseline for Predicting Out of Vocabulary Word Embeddings

COLING 2022main

Social media is the ultimate challenge for many natural language processing tools. The constant emergence of linguistic constructs challenge even the most sophisticated NLP tools. Predicting word embeddings for out of vocabulary words is one of those challenges. Word embedding models only include te…

Cited by 0SourcePDFScholar
2021

Segmentation of Tweets with URLs and its Applications to Sentiment Analysis

AAAI 2021technical

An important means for disseminating information in social media platforms is by including URLs that point to external sources in user posts. In Twitter, we estimate that about 21% of the daily stream of English-language tweets contain URLs. We notice that NLP tools make little attempt at understand…

Cited by 14SourcePDFScholar