← Search

Yuchong Xie

1 accepted papers

2026

From Similarity to Vulnerability: Key Collision Attack on LLM Semantic Caching

ICML 2026poster

Semantic caching has emerged as a pivotal technique for scaling LLM applications, widely adopted by major providers including AWS and Microsoft. By utilizing semantic embedding vectors as cache keys, this mechanism effectively minimizes latency and redundant computation for semantically similar quer…

Cited by 0SourceScholar