Fetching the paper…
Reading the bibliography…
We study social learning dynamics motivated by reviews on online platforms.
On the likelihood that one unknown probability exceeds another in view of the evidence of two samples
William R. Thompson · 1933
Earlier work this paper cites.
Etude critique de la notion de collectif
Jean Ville · 1939
Earlier work this paper cites.
Individual Choice Behavior: A Theoretical Analysis
R. Duncan Luce · 1959
Earlier work this paper cites.
Probability inequalities for sums of bounded random variables
Wassily Hoeffding · 1963
Earlier work this paper cites.
Reaching a Consensus
Morris H. DeGroot · 1974
Earlier work this paper cites.
Probability learning and sequence learning
Jerome L. Myers · 1976
Earlier work this paper cites.
Judgment under uncertainty: Heuristics and biases
Daniel Kahneman and Amos Tversky · 1982
Earlier work this paper cites.
Iterated least squares in multiperiod control
T.L. Lai and Herbert Robbins · 1982
Earlier work this paper cites.
Asymptotically efficient Adaptive Allocation Rules
Tze Leung Lai and Herbert Robbins · 1985
Earlier work this paper cites.
A simple model of herd behavior
Abhijit V. Banerjee · 1992
Earlier work this paper cites.
A Theory of Fads, Fashion, Custom, and Cultural Change as Informational Cascades
Sushil Bikhchandani, David Hirshleifer, and Ivo Welch · 1992
Earlier work this paper cites.
Sequential sales, learning, and cascades
Ivo Welch · 1992
Earlier work this paper cites.
The nonstochastic multiarmed bandit problem
Peter Auer, Nicolò Cesa-Bianchi, Yoav Freund, and Robert E. Schapire · 1995
Earlier work this paper cites.
Case-Based Decision Theory
Itzhak Gilboa and David Schmeidler · 1995
Earlier work this paper cites.
Learning from neighbours
Venkatesh Bala and Sanjeev Goyal · 1998
Earlier work this paper cites.
Strategic Experimentation
Patrick Bolton and Christopher Harris · 1999
Earlier work this paper cites.
Optimism and pessimism: Implications for theory, research, and practice
Edward C Chang, editor · 2000
Earlier work this paper cites.
Pathological outcomes of observational learning
Lones Smith and Peter Sørensen · 2000
Earlier work this paper cites.
An economist’s perspective on probability matching
Nir Vulkan · 2000
Earlier work this paper cites.
A survey of behavioral finance
Nicholas Barberis and Richard Thaler · 2003
Earlier work this paper cites.
Strategic Experimentation with Exponential Bandits
Godfrey Keller, Sven Rady, and Martin Cripps · 2005
Earlier work this paper cites.
The network structure of exploration and exploitation
David Lazer and Allan Friedman · 2007
Cited alongside, same era.
Optimism and economic choice
Manju Puri and David T. Robinson · 2007
Cited alongside, same era.
Regret Bounds and Minimax Policies under Partial Monitoring
J.Y. Audibert and S. Bubeck · 2009
Cited alongside, same era.
Multi-Armed Bandit Allocation Indices
John Gittins, Kevin Glazebrook, and Richard Weber · 2011
Cited alongside, same era.
Analysis of Thompson Sampling for the multi-armed bandit problem
Shipra Agrawal and Navin Goyal · 2012
Cited alongside, same era.
Regret Analysis of Stochastic and Nonstochastic Multi-armed Bandit Problems
Sébastien Bubeck and Nicolo Cesa-Bianchi · 2012
Cited alongside, same era.
Mostly exploration-free algorithms for contextual bandits
Hamsa Bastani, Mohsen Bayati, and Khashayar Khosravi · 2017
Later among the works it cites.
Learning, experimentation, and information design
Johannes Hörner and Andrzej Skrzypacz · 2017
Later among the works it cites.
Incentivizing exploration by heterogeneous users
Bangrui Chen, Peter I. Frazier, and David Kempe · 2018
Later among the works it cites.
Unrealistic expectations and misguided learning
Paul Heidhues, Botond Koszegi, and Philipp Strack · 2018
Later among the works it cites.
A smoothed analysis of the greedy algorithm for the linear contextual bandit problem
Sampath Kannan, Jamie Morgenstern, Aaron Roth, Bo Waggoner, and Zhiwei Steven Wu · 2018
Later among the works it cites.
On incomplete learning and certainty-equivalence control
N. Bora Keskin and Assaf Zeevi · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bayesian dynamic pricing policies: Learning and earning under a binary prior distribution
J. Michael Harrison, N. Bora Keskin, and Assaf Zeevi · 2012
Cited alongside, same era.
Thompson sampling: An asymptotically optimal finite-time analysis
Emilie Kaufmann, Nathaniel Korda, and Rémi Munos · 2012
Cited alongside, same era.
Near-optimal regret bounds for thompson sampling
Shipra Agrawal and Navin Goyal · 2013
Cited alongside, same era.
Recommender systems as mechanisms for social learning
Yeon-Koo Che and Johannes Hörner · 2013
Cited alongside, same era.
Implementing the “wisdom of the crowd”
Ilan Kremer, Yishay Mansour, and Motty Perry · 2013
Cited alongside, same era.
Simultaneously learning and optimizing using controlled variance pricing
Arnoud V. den Boer and Bert Zwart · 2014
Cited alongside, same era.
Greedy algorithm almost dominates in smoothed contextual bandits
Manish Raghavan, Aleksandrs Slivkins, Jennifer Wortman Vaughan, and Zhiwei Steven Wu · 2018
Later among the works it cites.
A tutorial on thompson sampling
Daniel Russo, Benjamin Van Roy, Abbas Kazerouni, Ian Osband, and Zheng Wen · 2018
Later among the works it cites.
Higher order concentration for functions of weakly dependent random variables
Friedrich Götze, Holger Sambale, and Arthur Sinulis · 2019
Later among the works it cites.
Introduction to multi-armed bandits
Aleksandrs Slivkins · 2019
Later among the works it cites.
Unreasonable effectiveness of greedy algorithms in multi-armed bandit with many arms
Mohsen Bayati, Nima Hamidi, Ramesh Johari, and Khashayar Khosravi · 2020
Later among the works it cites.
Incentivizing exploration with selective data disclosure
Nicole Immorlica, Jieming Mao, Aleksandrs Slivkins, and Steven Wu · 2020
Later among the works it cites.
Bandit Algorithms
Tor Lattimore and Csaba Szepesvári · 2020
Later among the works it cites.
On the non-asymptotic and sharp lower tail bounds of random variables
Anru R Zhang and Yuchen Zhou · 2020
Later among the works it cites.
Learning with Heterogeneous Misspecified Models: Characterization and Robustness
Aislinn Bohren and Daniel N. Hauser · 2021
Later among the works it cites.
Limit Points of Endogenous Misspecified Learning
Drew Fudenberg, Giacomo Lanzani, and Philipp Strack · 2021
Later among the works it cites.
Be greedy in multi-armed bandits, 2021
Matthieu Jedor, Jonathan Louëdec, and Vianney Perchet · 2021
Later among the works it cites.
The price of incentivizing exploration: A characterization via thompson sampling and sample complexity
Mark Sellke and Aleksandrs Slivkins · 2021
Later among the works it cites.
Dynamic Concern for Misspecification, 2023
Giacomo Lanzani · 2023
Closest in time.
Aleksandrs Slivkins · 2023
Closest in time.
Aleksandrs Slivkins, Yunzong Xu, and Shiliang Zuo · 2025
Closest in time.