Fetching the paper…
Reading the bibliography…
Effective network congestion control strategies are key to keeping the Internet (or any large computer network) operational.
On bayesian methods for seeking the extremum
J. Močkus · 1975
Earlier work this paper cites.
Markov Decision Processes: Discrete Stochastic Dynamic Programming
M. L. Puterman · 1994
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Overview of mini-batch gradient descent, 2012
G. Hinton, N. Srivastava, and K. Swersky · 2012
Earlier work this paper cites.
Tcp ex machina: Computer-generated congestion control
K. Winstein and H. Balakrishnan · 2013
Earlier work this paper cites.
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba · 2016
Earlier work this paper cites.
Automatic differentiation in PyTorch
A. Paszke, S. Gross, S. Chintala, G. Chanan, E. Yang, Z. DeVito, Z. Lin, A. Desmaison, L. Antiga, and A. Lerer · 2017
Cited alongside, same era.
Copa: Practical delay-based congestion control for the internet
V. Arun and H. Balakrishnan · 2018
Cited alongside, same era.
IMPALA: Scalable distributed deep-RL with importance weighted actor-learner architectures
L. Espeholt, H. Soyer, R. Munos, K. Simonyan, V. Mnih, T. Ward, Y. Doron, V. Firoiu, T. Harley, I. Dunning, S. Legg, and K. Kavukcuoglu · 2018
Cited alongside, same era.
Restructuring endpoint congestion control
A. Narayan, F. Cangialosi, D. Raghavan, P. Goyal, S. Narayana, R. Mittal, M. Alizadeh, and H. Balakrishnan · 2018
Cited alongside, same era.
Time limits in reinforcement learning, 2018
F. Pardo, A. Tavakoli, V. Levdik, and P. Kormushev · 2018
Cited alongside, same era.
Iroko: A framework to prototype reinforcement learning for data center traffic control
F. Ruffy, M. Przystupa, and I. Beschastnikh · 2018
Later among the works it cites.
Pantheon: the training ground for internet congestion-control research
F. Y. Yan, J. Ma, G. D. Hill, D. Raghavan, R. S. Wahby, P. Levis, and K. Winstein · 2018
Later among the works it cites.
A deep reinforcement learning perspective on internet congestion control
N. Jay, N. Rotman, B. Godfrey, M. Schapira, and A. Tamar · 2019
Closest in time.
TorchBeast: A PyTorch Platform for Distributed RL
H. Küttler, N. Nardelli, T. Lavril, M. Selvatici, V. Sivakumar, T. Rocktäschel, and E. Grefenstette · 2019
Closest in time.
Park: An open platform for learning augmented computer systems
H. Mao, P. Negi, A. Narayan, H. Wang, J. Yang, H. Wang, R. Marcus, R. Addanki, M. Khani, S. He, V. Nathan, F. Cangialosi, S. B. Venkatakrishnan, W.-H. Weng, S. Han, T. Kraska, and M. Alizadeh · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Closest in time.