← Search

Stefan Schnake

2 accepted papers

2026

Tucker Attention: A generalization of approximate attention mechanisms

ICML 2026poster

The pursuit of reducing the memory footprint of the self-attention mechanism in multi-headed self attention (MHA) spawned a rich portfolio of methods, e.g., group-query attention (GQA) and multi-head latent attention (MLA). The methods leverage specialized low-rank factorizations across embedding di…

Cited by 0SourceScholar
2025

Dynamical Low-Rank Compression of Neural Networks with Robustness under Adversarial Attacks

NeurIPS 2025oral

Deployment of neural networks on resource-constrained devices demands models that are both compact and robust to adversarial inputs. However, compression and adversarial robustness often conflict. In this work, we introduce a dynamical low-rank training scheme enhanced with a novel spectral regulari…

Cited by 0SourceScholar