Positional Differential Encoding for Distributed Learning
A growing amount of available data and computational power makes training neural networks over a network of devices, and distribution optimization in general, more realizable. As a consequence, efficient communication becomes more important and can often be the bottleneck of such algorithms. Traditi…