Fetching the paper…
Reading the bibliography…
In this paper, we utilize results from convex analysis and monotone operator theory to derive additional properties of the softmax function that have not yet been covered in the existing literature.
R. Luce, Individual Choice Behavior: A Theoretical Analysis. NY, Wiley, 1959
1959
Earlier work this paper cites.
J. Smith and G. Price, “The Logic of Animal Conflict”, Nature, vol. 246, no. 5427, pp. 15-18, 1973
1973
Earlier work this paper cites.
J. Baillon and G. Haddad, “Quelques propriétés des opérateurs angle-bornés etn-cycliquement monotones”, Israel Journal of Mathematics, vol. 26, no. 2, pp. 137-150, 1977
1977
Earlier work this paper cites.
J. Bridle, “Probabilistic Interpretation of Feedforward Classification Network Outputs, with Relationships to Statistical Pattern Recognition”, Neurocomputing: Algorithms, Architectures and Applications, F. Soulie and J. Herault, eds., pp. 227-236, 1990
1990
Earlier work this paper cites.
R. McKelvey and T. Palfrey, Quantal response equilibria for normal form games, 1st ed. Pasadena, Calif.: Division of the Humanities and Social Sciences, California Institute of Technology, 1994
1994
Earlier work this paper cites.
I. M. Elfadel and J. L. Wyatt Jr., “The softmax nonlinearity: Derivation using statistical mechanics and useful properties as a multiterminal analog circuit element”, In Advances in Neural Information Processing Systems 6, J. Cowan, G. Tesauro, and C. L. Giles, Eds. San Mateo, CA: Morgan Kaufmann, 1994, pp. 882-887
1994
Earlier work this paper cites.
J. Weibull, “ Evolutionary Game Theory”. MIT Press, Cambridge, 1995
1995
Earlier work this paper cites.
A. L. Yuille and D. Geiger, “Winner-Take-All Mechanisms”, In The Handbook of Brain Theory and Neural Networks, Ed. M. Arbib, MIT Press, 1995
1995
Earlier work this paper cites.
I. M. Elfadel, “Convex Potentials and their Conjugates in Analog Mean-Field Optimization”, Neural Computation, vol. 7, no. 5, pp. 1079-1104, 1995
1995
Earlier work this paper cites.
R. Sutton and A. Barto, Reinforcement Learning: An Introduction. Cambridge, MA, USA: MIT Press, 1998
1998
Earlier work this paper cites.
J. Hofbauer, K. Sigmund, “Evolutionary Games and Population Dynamics”, Cambridge University Press, 1998
1998
Earlier work this paper cites.
R. T. Rockafellar and R. J.-B. Wets, Variational Analysis. Berlin: Springer-Verlag, 1998
1998
Earlier work this paper cites.
A. Rangarajan, “Self-annealing and self-annihilation: unifying deterministic annealing and relaxation labeling”, Pattern Recognition, vol. 33, no. 4, pp. 635-649, 2000
2000
Earlier work this paper cites.
J. B. Hiriart-Urruty and C. Lemaréchal: Fundamentals of Convex Analysis. Springer–Verlag, Berlin 2001
2001
Earlier work this paper cites.
R. Zunino, P. Gastaldo, “Analog implementation of the softmax function”, In IEEE International Symposium on Circuits and Systems, vol 2, pp II-117, 2002
2002
Earlier work this paper cites.
H. K. Khalil, Nonlinear Systems, 3rd ed., Upper Siddle River, NJ: Prentice-Hall, 2002
2002
Earlier work this paper cites.
E. Hopkins, “Two Competing Models of How People Learn in Games”, Econometrica, vol. 70, no. 6, pp. 2141-2166, 2002
2002
Earlier work this paper cites.
F. Facchinei and J.-S. Pang, Finite-dimensional Variational Inequalities and Complementarity Problems. Vol. I, Springer Series in Operations Research, Springer-Verlag, New York, 2003
2003
Earlier work this paper cites.
A. Beck and M. Teboulle, “Mirror descent and nonlinear projected subgradient methods for convex optimization”, Operations Research Letters, vol. 31, no. 3, pp. 167-175, 2003
2003
Earlier work this paper cites.
Y. Sato and J. Crutchfield, “Coupled Replicator Equations for the dynamics of learning in multiagent systems”, Physical Review E, vol. 67, no. 1, 2003
2003
Earlier work this paper cites.
K. Tuyls, K. Verbeeck and T. Lenaerts, “A selection-mutation model for Q-learning in multi-agent systems”, in Proc. of the 2nd Int. Joint Conf. on Autonomous Agents and Multi-Agent Systems (AAMAS), pp. 693-700, 2003
2003
Earlier work this paper cites.
S. Boyd and L. Vandenberghe, Convex optimization, 1st ed. Cambridge, UK: Cambridge University Press, 2004
2004
Cited alongside, same era.
Y. Nesterov, Introductory Lectures on Convex Optimization: A Basic Course. Norwell, MA: Kluwer, 2004
2004
Cited alongside, same era.
F. Alvarez, J. Bolte and O. Brahic, “Hessian Riemannian Gradient Flows in Convex Programming”, SIAM Journal on Control and Optimization, vol. 43, no. 2, pp. 477-501, 2004
2004
Cited alongside, same era.
J. Hofbauer and E. Hopkins, “Learning in perturbed asymmetric games”, Games Economic Behav., vol. 52, pp. 133-152, 2005
2005
Cited alongside, same era.
D. Leslie and E. Collins, “Individual Q-Learning in Normal Form Games”, SIAM Journal on Control and Optimization, vol. 44, no. 2, pp. 495-514, 2005
2005
Cited alongside, same era.
A. Kianercy and A. Galstyan, “Dynamics of Boltzmann Q-learning in two-player two-action games”, Phys. Rev. E, vol. 85, no. 4, pp. 1145-1154, 2012
2012
Later among the works it cites.
P. Mertikopoulos, E. V. Belmega, and A. L. Moustakas, “Matrix exponential learning: Distributed optimization in MIMO systems”, in ISIT’12: Proceedings of the 2012 IEEE International Symposium on Information Theory, 2012, pp. 3028-3032
2012
Later among the works it cites.
G. Weiss, Multiagent Systems, 2nd ed. Cambridge, MA, USA: MIT Press, 2013
2013
Later among the works it cites.
R. Laraki and P. Mertikopoulos, “Higher order game dynamics”, Journal of Economic Theory, vol. 148, no. 6, pp. 2666-2695, 2013
2013
Later among the works it cites.
E. Alpaydin, Introduction to Machine Learning, 3rd ed. The MIT Press, 2014, p. 264
2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. M. Bishop, Pattern Recognition and Machine Learning. Secaucus, NJ, USA: Springer, 2006
2006
Cited alongside, same era.
N. Daw, J. O’Doherty, P. Dayan, B. Seymour and R. Dolan, “Cortical substrates for exploratory decisions in humans”, Nature, vol. 441, no. 7095, pp. 876-879, 2006
2006
Cited alongside, same era.
D. Lee, “Neuroeconomics: Best to go with what you know?”, Nature, vol. 441, no. 7095, pp. 822-823, 2006
2006
Cited alongside, same era.
J. D. Cohen, S. M. McClure, and A. J. Yu, “Should I stay or should I go? How the human brain manages the trade-off between exploitation and exploration”, Philosph. Trans. Roy. Soc. B: Bio. Sci., vol. 362, no. 1481, pp. 933-942, 2007
2007
Cited alongside, same era.
D. Koulouriotis and A. Xanthopoulos, “Reinforcement learning and evolutionary algorithms for non-stationary multi-armed bandit problems”, Applied Mathematics and Computation, vol. 196, no. 2, pp. 913-922, 2008
2008
Cited alongside, same era.
W. Sandholm, E. Dokumacı and R. Lahkar, “The projection dynamic and the replicator dynamic”, Games and Economic Behavior, vol. 64, no. 2, pp. 666-683, 2008
2008
Cited alongside, same era.
J. Hofbauer and W. H. Sandholm, “Stable games and their dynamics”, Journal of Economic Theory, vol. 144, no. 4, pp. 1665-1693, 2009
2009
Cited alongside, same era.
H. Young and S. Zamir, Handbook of Game Theory, Volume 4, 1st ed. Amsterdam: Elsevier, North-Holland, 2015
2015
Later among the works it cites.
D. Bloembergen, K. Tuyls, D. Hennes, and M. Kaisers, “Evolutionary dynamics of multi-agent learning: A survey”, J. Artif. Intell. Res., vol. 53, no. 1, pp. 659-697, May 2015
2015
Later among the works it cites.
P. Coucheney, B. Gaujal and P. Mertikopoulos, “Penalty-Regulated Dynamics and Robust Learning Procedures in Games”, Mathematics of Operations Research, vol. 40, no. 3, pp. 611-633, 2015
2015
Later among the works it cites.
J. Peypouquet. Convex optimization in normed spaces: theory, methods and examples. Springer, 2015
2015
Later among the works it cites.
P. Bossaerts and C. Murawski, “From behavioural economics to neuroeconomics to decision neuroscience: the ascent of biology in research on human decision making”, Current Opinion in Behavioral Sciences, vol. 5, pp. 37-42, 2015
2015
Later among the works it cites.
J. Goeree, C. Holt and T. Palfrey, Quantal response equilibrium: A Stochastic Theory of Games. Princeton University Press, 2016
2016
Later among the works it cites.
I. Goodfellow, Y. Bengio, and A. Courville, Deep Learning. Cambridge, MA, USA: MIT Press, 2016
2016
Later among the works it cites.
P. Mertikopoulos and W. Sandholm, “Learning in Games via Reinforcement and Regularization”, Mathematics of Operations Research, vol. 41, no. 4, pp. 1297-1324, 2016
2016
Later among the works it cites.
T. Genewein and D. A. Braun, “Bio-inspired feedback-circuit implementation of discrete, free energy optimizing, winner-take-all computations”, Biological, vol. 110, no. 2, pp. 135-150, Jun. 2016
2016
Later among the works it cites.
T. Michalis, “One-vs-each approximation to softmax for scalable estimation of probabilities”, In Advances in Neural Information Processing Systems 29, pp. 4161-4169. 2016
2016
Later among the works it cites.
P. Reverdy and N. Leonard, “Parameter Estimation in Softmax Decision-Making Models With Linear Objective Functions”, IEEE Transactions on Automation Science and Engineering, vol. 13, no. 1, pp. 54-67, 2016
2016
Later among the works it cites.
2016
Later among the works it cites.
E. Hazan, “Introduction to Online Convex Optimization”, Foundations and Trends® in Optimization, vol. 2, no. 3-4, pp. 157-325, 2016
2016
Later among the works it cites.
2016
Later among the works it cites.