← Search

Artin Spiridonoff

1 accepted papers

2021

Communication-efficient SGD: From Local SGD to One-Shot Averaging

NeurIPS 2021poster

We consider speeding up stochastic gradient descent (SGD) by parallelizing it across multiple workers. We assume the same data set is shared among $N$ workers, who can take SGD steps and coordinate with a central server. While it is possible to obtain a linear reduction in the variance by averaging…

Cited by 29SourcePDFScholar