← Search

Hanxiang Hao

1 accepted papers

2023

UPSCALE: Unconstrained Channel Pruning

ICML 2023poster

As neural networks grow in size and complexity, inference speeds decline. To combat this, one of the most effective compression techniques -- channel pruning -- removes channels from weights. However, for multi-branch segments of a model, channel removal can introduce inference-time memory copies. I…