← Search

Rajit Rajpal

1 accepted papers

2026

Adaptive Momentum and Nonlinear Damping for Neural Network Training

ICML 2026poster

Momentum Stochastic Gradient Descent (mSGD) relies on a fixed momentum coefficient shared across all parameters, failing to account for the heterogeneous structure of modern loss landscapes. In this work, we adopt a continuous-time formulation to introduce individual, adaptive momentum coefficients …

Cited by 0SourceScholar