2025
A unified framework for establishing the universal approximation of transformer-type architectures
NeurIPS 2025poster
We investigate the universal approximation property (UAP) of transformer-type architectures, providing a unified theoretical framework that extends prior results on residual networks to models incorporating attention mechanisms. Our work identifies token distinguishability as a fundamental requireme…