← Search

Karim Ghonim

4 accepted papers

2025

Concept-pedia: a Wide-coverage Semantically-annotated Multimodal Dataset

EMNLP 2025

Vision-language Models (VLMs), such as CLIP and SigLIP, have become the de facto standard for multimodal tasks, serving as essential building blocks for recent Multimodal Large Language Models, including LLaVA and PaliGemma. However, current evaluations for VLMs remain heavily anchored to ImageNet.

Cited by 0SourcePDFScholar
2025

RAED: Retrieval-Augmented Entity Description Generation for Emerging Entity Linking and Disambiguation

EMNLP 2025

Entity Linking and Entity Disambiguation systems aim to link entity mentions to their corresponding entries, typically represented by descriptions within a predefined, static knowledge base. Current models assume that these knowledge bases are complete and up-to-date, rendering them incapable of han

2024

FENICE: Factuality Evaluation of summarization based on Natural language Inference and Claim Extraction

ACL 2024findings

Recent advancements in text summarization, particularly with the advent of Large Language Models (LLMs), have shown remarkable performance. However, a notable challenge persists as a substantial number of automatically-generated summaries exhibit factual inconsistencies, such as hallucinations. In r…

2024

Mitigating Data Scarcity in Semantic Parsing across Languages with the Multilingual Semantic Layer and its Dataset

ACL 2024findings

Data scarcity is a prevalent challenge in the era of Large Language Models (LLMs). The insatiable hunger of LLMs for large corpora becomes even more pronounced when dealing with non-English and low-resource languages. The issue is particularly exacerbated in Semantic Parsing (SP), i.e. the task of c…