← Search

Hoda Eldardiry

7 accepted papers

2025

MMPlanner: Zero-Shot Multimodal Procedural Planning with Chain-of-Thought Object State Reasoning

EMNLP 2025

Multimodal Procedural Planning (MPP) aims to generate step-by-step instructions that combine text and images, with the central challenge of preserving object-state consistency across modalities while producing informative plans. Existing approaches often leverage large language models (LLMs) to refi

Cited by 0SourcePDFScholar
2025

Sci-LoRA: Mixture of Scientific LoRAs for Cross-Domain Lay Paraphrasing

ACL 2025finding

Lay paraphrasing aims to make scientific information accessible to audiences without technical backgrounds. However, most existing studies focus on a single domain, such as biomedicine. With the rise of interdisciplinary research, it is increasingly necessary to comprehend knowledge spanning multipl…

2025

VTechAGP: An Academic-to-General-Audience Text Paraphrase Dataset and Benchmark Models

NAACL 2025long

Existing text simplification or paraphrase datasets mainly focus on sentence-level text generation in a general domain. These datasets are typically developed without using domain knowledge. In this paper, we release a novel dataset, VTechAGP, which is the first academic-to-general-audience text par…

2025

Visual Zero-Shot E-Commerce Product Attribute Value Extraction

NAACL 2025industry

Existing zero-shot product attribute value (aspect) extraction approaches in e-Commerce industry rely on uni-modal or multi-modal models, where the sellers are asked to provide detailed textual inputs (product descriptions) for the products. However, manually providing (typing) the product descripti…

2024

Prompt-based Zero-shot Relation Extraction with Semantic Knowledge Augmentation

COLING 2024main

In relation triplet extraction (RTE), recognizing unseen relations for which there are no training instances is a challenging task. Efforts have been made to recognize unseen relations based on question-answering models or relation descriptions. However, these approaches miss the semantic informatio…