← Search

Weitao Li

2 accepted papers

2025

Efficient Dynamic Clustering-Based Document Compression for Retrieval-Augmented-Generation

EMNLP 2025

Retrieval-Augmented Generation (RAG) has emerged as a widely adopted approach for knowledge injection during large language model (LLM) inference in recent years. However, due to their limited ability to exploit fine-grained inter-document relationships, current RAG implementations face challenges i