← Search

Santiago Akle Serrano

1 accepted papers

2020

Stochastic Weight Averaging in Parallel: Large-Batch Training That Generalizes Well

ICLR 2020poster

We propose Stochastic Weight Averaging in Parallel (SWAP), an algorithm to accelerate DNN training. Our algorithm uses large mini-batches to compute an approximate solution quickly and then refines it by averaging the weights of multiple models computed independently and in parallel. The resulting m…

Cited by 65SourceScholar