← Search

Zhenxin Huang

1 accepted papers

2026

ArborKV: Structure-Aware KV Cache Management for Scaling Tree-based LLM Reasoning

ICML 2026poster

Recent progress in LLM reasoning has increasingly shifted from single-pass generation to explicit search over intermediate reasoning states. Tree-of-Thoughts (ToT) organizes inference to tree-structured search with branching and backtracking, but it substantially amplifies the key--value (KV) cache:…

Cited by 0SourceScholar