← Search

Alejandro Jaimes

15 accepted papers

2025

CEHA: A Dataset of Conflict Events in the Horn of Africa

COLING 2025main

Natural Language Processing (NLP) of news articles can play an important role in understanding the dynamics and causes of violent conflict. Despite the availability of datasets categorizing various conflict events, the existing labels often do not cover all of the fine-grained violent conflict event…

2025

Explain then Rank: Scale Calibration of Neural Rankers Using Natural Language Explanations from LLMs

ACL 2025finding

In search settings, calibrating the scores during the ranking process to quantities such as click-through rates or relevance levels enhances a system’s usefulness and trustworthiness for downstream users. While previous research has improved this notion of calibration for low complexity learning-to-…

2025

Uchaguzi-2022: A Dataset of Citizen Reports on the 2022 Kenyan Election

COLING 2025main

Online reporting platforms have enabled citizens around the world to collectively share their opinions and report in real time on events impacting their local communities. Systematically organizing (e.g., categorizing by attributes) and geotagging large amounts of crowdsourced information is crucial…

Cited by 0SourcePDFScholar
2024

HumVI: A Multilingual Dataset for Detecting Violent Incidents Impacting Humanitarian Aid

EMNLP 2024finding

Humanitarian organizations can enhance their effectiveness by analyzing data to discover trends, gather aggregated insights, manage their security risks, support decision-making, and inform advocacy and funding proposals. However, data about violent incidents with direct impact and relevance for hum…

2023

A New Task and Dataset on Detecting Attacks on Human Rights Defenders

ACL 2023findings

The ability to conduct retrospective analyses of attacks on human rights defenders over time and by location is important for humanitarian organizations to better understand historical or ongoing human rights violations and thus better manage the global impact of such events. We hypothesize that NLP…

2023

BUMP: A Benchmark of Unfaithful Minimal Pairs for Meta-Evaluation of Faithfulness Metrics

ACL 2023long

The proliferation of automatic faithfulness metrics for summarization has produced a need for benchmarks to evaluate them. While existing benchmarks measure the correlation with human judgements of faithfulness on model-generated summaries, they are insufficient for diagnosing whether metrics are: 1…

2023

Harnessing the power of LLMs: Evaluating human-AI text co-creation through the lens of news headline generation

EMNLP 2023long findings

To explore how humans can best leverage LLMs for writing and how interacting with these models affects feelings of ownership and trust in the writing process, we compared common human-AI interaction types (e.g., guiding system, selecting from system outputs, post-editing outputs) in the context of L…

Cited by 0SourcecodeScholar
2022

CrisisLTLSum: A Benchmark for Local Crisis Event Timeline Extraction and Summarization

EMNLP 2022finding

Social media has increasingly played a key role in emergency response: first responders can use public posts to better react to ongoing crisis events and deploy the necessary resources where they are most needed. Timeline extraction and abstractive summarization are critical technical tasks to lever…

2022

XLTime: A Cross-Lingual Knowledge Transfer Framework for Temporal Expression Extraction

NAACL 2022findings

Temporal Expression Extraction (TEE) is essential for understanding time in natural language. It has applications in Natural Language Processing (NLP) tasks such as question answering, information retrieval, and causal inference. To date, work in this area has mostly focused on English as there is a…

2021

Journalistic Guidelines Aware News Image Captioning

EMNLP 2021main

The task of news article image captioning aims to generate descriptive and informative captions for news article images. Unlike conventional image captions that simply describe the content of the image in general terms, news image captions follow journalistic guidelines and rely heavily on named ent…

2020

Multimodal Categorization of Crisis Events in Social Media

CVPR 2020poster

Recent developments in image classification and natural language processing, coupled with the rapid growth in social media usage, have enabled fundamental advances in detecting breaking events around the world in real-time. Emergency response is one such area that stands to gain from these advances.…

Cited by 138PDFScholar
2016

TGIF: A New Dataset and Benchmark on Animated GIF Description

CVPR 2016spotlight

With the recent popularity of animated GIFs on social media, there is need for ways to index them with rich metadata. To advance research on animated GIF understanding, we collected a new dataset, Tumblr GIF (TGIF), with 100K animated GIFs from Tumblr and 120K natural language descriptions obtained…

Cited by 330PDFcodeScholar