Fetching the paper…
Reading the bibliography…
We investigate a classification problem using multiple mobile agents capable of collecting (partial) pose-dependent observations of an unknown environment.
C. Connolly, “The determination of next best views,” in Proceedings. 1985 IEEE International Conference on Robotics and Automation , vol. 2. IEEE, 1985, pp. 432–435
1985
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-based learning applied to document recognition,” Proceedings of the IEEE , vol. 86, no. 11, pp. 2278–2324, 1998
1998
Earlier work this paper cites.
R. S. Sutton, D. A. McAllester, S. P. Singh, and Y. Mansour, “Policy gradient methods for reinforcement learning with function approximation,” in Advances in neural information processing systems , 2000, pp. 1057–1063
2000
Earlier work this paper cites.
L. v. d. Maaten and G. Hinton, “Visualizing data using t-sne,” Journal of machine learning research , vol. 9, no. Nov, pp. 2579–2605, 2008
2008
Earlier work this paper cites.
A. D. Domínguez-García and C. N. Hadjicostis, “Distributed strategies for average consensus in directed graphs,” in 2011 50th IEEE Conference on Decision and Control and European Control Conference . IEEE, 2011, pp. 2124–2129
2011
Earlier work this paper cites.
L. Yushi, J. Fei, and Y. Hui, “Study on application modes of military internet of things (miot),” in 2012 IEEE International Conference on Computer Science and Automation Engineering (CSAE) , vol. 3. IEEE, 2012, pp. 630–634
2012
Earlier work this paper cites.
D. Silver, L. Newnham, D. Barker, S. Weller, and J. McFall, “Concurrent reinforcement learning from customer interactions,” in International Conference on Machine Learning , 2013, pp. 924–932
2013
Earlier work this paper cites.
R. Johnson and T. Zhang, “Accelerating stochastic gradient descent using predictive variance reduction,” in Advances in neural information processing systems , 2013, pp. 315–323
2013
Earlier work this paper cites.
M. A. K. Bahrin, M. F. Othman, N. H. N. Azli, and M. F. Talib, “Industry 4.0: A review on industrial automation and robotic,” Jurnal Teknologi , vol. 78, no. 6-13, 2016
2016
Cited alongside, same era.
2016
Cited alongside, same era.
J. Foerster, I. A. Assael, N. de Freitas, and S. Whiteson, “Learning to communicate with deep multi-agent reinforcement learning,” in Advances in Neural Information Processing Systems , 2016, pp. 2137–2145
2016
Cited alongside, same era.
I. Goodfellow, Y. Bengio, and A. Courville, Deep learning . MIT press, 2016
2016
Cited alongside, same era.
A. Paszke, S. Gross, S. Chintala, G. Chanan, E. Yang, Z. DeVito, Z. Lin, A. Desmaison, L. Antiga, and A. Lerer, “Automatic differentiation in pytorch,” in NIPS-W , 2017
2017
Later among the works it cites.
2018
Later among the works it cites.
A. Khan, C. Zhang, V. Kumar, and A. Ribeiro, “Collaborative multiagent reinforcement learning in homogeneous swarms,” 2018
2018
Later among the works it cites.
J. N. Foerster, G. Farquhar, T. Afouras, N. Nardelli, and S. Whiteson, “Counterfactual multi-agent policy gradients,” in Thirty-Second AAAI Conference on Artificial Intelligence , 2018
2018
Later among the works it cites.
M. Fazel, R. Ge, S. M. Kakade, and M. Mesbahi, “Global convergence of policy gradient methods for linearized control problems,” 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
C. Bhatt, N. Dey, and A. S. Ashour, Internet of things and big data technologies for next generation healthcare . Springer, 2017, vol. 23
2017
Cited alongside, same era.
R. Lowe, Y. Wu, A. Tamar, J. Harb, O. P. Abbeel, and I. Mordatch, “Multi-agent actor-critic for mixed cooperative-competitive environments,” in Advances in Neural Information Processing Systems , 2017, pp. 6379–6390
2017
Cited alongside, same era.
K. Bonawitz, V. Ivanov, B. Kreuter, A. Marcedone, H. B. McMahan, S. Patel, D. Ramage, A. Segal, and K. Seth, “Practical secure aggregation for privacy-preserving machine learning,” in Proceedings of the 2017 ACM SIGSAC Conference on Computer and Communications Security . ACM, 2017, pp. 1175–1191
2017
Cited alongside, same era.
T. Nguyen and S. Mukhopadhyay, “Selectively decentralized q-learning,” in 2017 IEEE International Conference on Systems, Man, and Cybernetics (SMC) . IEEE, 2017, pp. 328–333
2017
Cited alongside, same era.
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.