EMNLP 20250 citations

Topic-Guided Reinforcement Learning with LLMs for Enhancing Multi-Document Summarization

Chuyuan Li, Austin Xu, Shafiq Joty, Giuseppe Carenini

Abstract

A key challenge in Multi-Document Summarization (MDS) is effectively integrating information from multiple sources while maintaining coherence and topical relevance. While Large Language Models (LLMs) have shown impressive results in single-document summarization, their performance on MDS still leaves room for improvement. In this paper, we propose a topic-guided reinforcement learning approach to improve content selection in MDS. We first show that explicitly prompting models with topic labels enhances the informativeness. Building on this insight, we propose a novel topic reward within the Group Relative Policy Optimization (GRPO) framework to measure topic alignment between the generated summary and source documents. Experimental results on the Multi-News and Multi-XScience datasets demonstrate that our method consistently outperforms strong baselines, highlighting the effectiveness of leveraging topical cues in MDS.

BibTeX
@inproceedings{emnlp2025_topicguidedreinf,
  title = {Topic-Guided Reinforcement Learning with LLMs for Enhancing Multi-Document Summarization},
  author = {Chuyuan Li and Austin Xu and Shafiq Joty and Giuseppe Carenini},
  booktitle = {EMNLP 2025},
  year = {2025}
}
Topic-Guided Reinforcement Learning with LLMs for Enhancing Multi-Document Summarization · EMNLP 2025