Fetching the paper…
Reading the bibliography…
Compressed communication, in the form of sparsification or quantization of stochastic gradients, is employed to reduce communication costs in distributed data-parallel training of deep neural networks.
Nothing clear enough to list yet.
https://pytorch.org/
PyTorch
Cited in the paper.
https://tensorflow.org/
TensorFlow
Cited in the paper.
Nothing clear enough to list yet.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…