2025
GraphKV: Breaking the Static Selection Paradigm with Graph-Based KV Cache Eviction
EMNLP 2025
Efficient Key-Value (KV) cache management is essential for processing long text sequences in large language models (LLMs), where memory constraints often limit performance. Conventional KV eviction strategies, such as top-k selection based on attention scores, depend on static heuristics that fail t