← Search

Renhai Chen

3 accepted papers

2025

HATA: Trainable and Hardware-Efficient Hash-Aware Top-k Attention for Scalable Large Model Inference

ACL 2025finding

Large Language Models (LLMs) have emerged as a pivotal research area, yet the attention module remains a critical bottleneck in LLM inference, even with techniques like KVCache to mitigate redundant computations. While various top-k attention mechanisms have been proposed to accelerate LLM inference…

2025

Relaxing Distillation Constraints for Improved New Class Learning in Continual Semantic Segmentation

ICASSP 2025accepted

Continual Semantic Segmentation (CSS) aims to continuously learn new classes while mitigating catastrophic forgetting. Existing CSS methods primarily address this challenge through knowledge distillation. While they focus on maintaining stability for old classes, this emphasis often restricts plasti…

Cited by 0SourceScholar
2018

A Reliable Video Storage Architecture in Hybrid SLC/MLC Nand Flash

ICASSP 2018accepted

In this paper, we propose a reliable video storage architecture in hybrid SLC/MLC storage systems. In this architecture, the video stream is reconstructed as the key cluster and the non-key cluster according to the importance of video restoration. The key cluster is stored in SLC blocks to ensure th…

Cited by 0SourceScholar