Fetching the paper…
Reading the bibliography…
In order to mitigate the high communication cost in distributed and federated learning, various vector compression schemes, such as quantization, sparsification and dithering, have become very popular.
PowerSGD: Practical Low-Rank Gradient Compression for Distributed Optimization
\NAT@biblabelnum · 1905
Earlier work this paper cites.
Über den anschaulichen Inhalt der quantentheoretischen Kinematik und Mechanik
\NAT@biblabelnum · 1927
Earlier work this paper cites.
Theory of Communication
\NAT@biblabelnum · 1946
Earlier work this paper cites.
Television by pulse code modulation
\NAT@biblabelnum · 1951
Earlier work this paper cites.
Picture coding using pseudo-random noise
\NAT@biblabelnum · 1962
Earlier work this paper cites.
Diameters of some finite-dimensional sets and classes of smooth functions
\NAT@biblabelnum · 1977
Earlier work this paper cites.
The Uncertainty Principle in Harmonic Analysis
\NAT@biblabelnum · 1994
Earlier work this paper cites.
Constructive approximation of a ball by polytopes
\NAT@biblabelnum · 1994
Earlier work this paper cites.
Concentration property on probability spaces
\NAT@biblabelnum · 2000
Earlier work this paper cites.
The Concentration of Measure Phenomenon
\NAT@biblabelnum · 2001
Earlier work this paper cites.
A note on approximation of a ball by polytopes
\NAT@biblabelnum · 2004
Earlier work this paper cites.
Decoding by linear programming
\NAT@biblabelnum · 2005
Earlier work this paper cites.
Near-optimal signal recovery from random projections and universal encoding strategies
\NAT@biblabelnum · 2006
Earlier work this paper cites.
Elements of Information Theory (Wiley Series in Telecommunications and Signal Processing)
\NAT@biblabelnum · 2006
Earlier work this paper cites.
Uncertainty Principles and Vector Quantization
\NAT@biblabelnum · 2010
Earlier work this paper cites.
Scaling up machine learning: Parallel and distributed approaches
\NAT@biblabelnum · 2011
Cited alongside, same era.
Deep learning in neural networks: An overview
\NAT@biblabelnum · 2015
Cited alongside, same era.
Federated learning: strategies for improving communication efficiency
\NAT@biblabelnum · 2016
Cited alongside, same era.
QSGD: Communication-Efficient SGD via Gradient Quantization and Encoding
\NAT@biblabelnum · 2017
Cited alongside, same era.
Accurate, Large Minibatch SGD: Training ImageNet in 1 Hour
\NAT@biblabelnum · 2017
Cited alongside, same era.
Communication-efficient learning of deep networks from decentralized data
\NAT@biblabelnum · 2017
Cited alongside, same era.
Atomo: Communication-efficient learning via atomic sparsification
\NAT@biblabelnum · 2018
Later among the works it cites.
Gradient sparsification for communication-efficient distributed optimization
\NAT@biblabelnum · 2018
Later among the works it cites.
signSGD with majority vote is communication efficient and fault tolerant
\NAT@biblabelnum · 2019
Later among the works it cites.
Expanding the Reach of Federated Learning by Reducing Client Resource Requirements
\NAT@biblabelnum · 2019
Later among the works it cites.
Natural Compression for Distributed Deep Learning
\NAT@biblabelnum · 2019
Later among the works it cites.
Stochastic distributed learning with gradient quantization and variance reduction
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Terngrad: Ternary gradients to reduce communication in distributed deep learning
\NAT@biblabelnum · 2017
Cited alongside, same era.
ZipML: Training linear models with end-to-end low precision, and a little bit of deep learning
\NAT@biblabelnum · 2017
Cited alongside, same era.
The convergence of sparsified gradient methods
\NAT@biblabelnum · 2018
Cited alongside, same era.
signSGD: Compressed Optimisation for Non-Convex Problems
\NAT@biblabelnum · 2018
Cited alongside, same era.
Distributed learning with compressed gradients
\NAT@biblabelnum · 2018
Cited alongside, same era.
Randomized distributed mean estimation: accuracy vs communication
\NAT@biblabelnum · 2018
Cited alongside, same era.
\NAT@biblabelnum · 2019
Later among the works it cites.
SCAFFOLD: Stochastic Controlled Averaging for On-Device Federated Learning
\NAT@biblabelnum · 2019
Later among the works it cites.
Error Feedback Fixes SignSGD and other Gradient Compression Schemes
\NAT@biblabelnum · 2019
Later among the works it cites.
Federated learning: challenges, methods, and future directions
\NAT@biblabelnum · 2019
Later among the works it cites.
signSGD via zeroth-order oracle
\NAT@biblabelnum · 2019
Later among the works it cites.
On stochastic sign descent methods
\NAT@biblabelnum · 2019
Later among the works it cites.
DoubleSqueeze
\NAT@biblabelnum · 2019
Later among the works it cites.
Fast and Faster Convergence of SGD for Over-Parameterized Models and an Accelerated Perceptron
\NAT@biblabelnum · 2019
Later among the works it cites.
Tighter theory for local SGD on identical and heterogeneous data
\NAT@biblabelnum · 2020
Closest in time.