← Search

Jaewon Cheon

1 accepted papers

2025

COUNTDOWN: Contextually Sparse Activation Filtering Out Unnecessary Weights in Down Projection

EMNLP 2025

The growing size of large language models has created significant computational inefficiencies. To address this challenge, sparse activation selectively deactivates non-essential parameters during inference, reducing computational costs in FFNN layers. While existing methods focus on non-linear gati