← Search

Shaochen (Henry) Zhong

2 accepted papers

2026

FAFO: Lossy KV Cache Compression for Lossless Inference Acceleration via Draftless Fumble Decoding

ICML 2026poster

Lossy KV cache compression is a well-explored subfield of machine learning efficiency, with improved latency being one of its major gains. However, lossy compression techniques can fumble from time to time, exhibiting various — and often catastrophic — failure patterns that are not only difficult to…

Cited by 0SourceScholar
2026

Position: Want Better ML Reviews? Stop Asking Nicely and Start Incentivizing with a Credit System

ICML 2026poster

With soaring submission counts, stricter reciprocal review policies, widespread adoption of platforms like OpenReview, and without the offsetting pressure of publication fees, the machine learning (ML) community has one of the largest scholarly presences among all scientific fields. And yet, **almos…

Cited by 0SourceScholar