Fetching the paper…
Reading the bibliography…
We present new excess risk bounds for general unbounded loss functions including log loss and squared loss, where the distribution of the losses may be heavy-tailed.
On information and sufficiency
Solomon Kullback and Richard A. Leibler · 1951
Earlier work this paper cites.
Convex Analysis
R. Tyrrell Rockafellar · 1970
Earlier work this paper cites.
Problem complexity and method efficiency in optimization
Arkadii Nemirovskii and David Borisovich Yudin · 1983
Earlier work this paper cites.
Generalized Linear Models
Peter McCullagh and John Nelder · 1989
Earlier work this paper cites.
Stochastic Complexity in Statistical Inquiry
Jorma Rissanen · 1989
Earlier work this paper cites.
Aggregating strategies
Vladimir Vovk · 1990
Earlier work this paper cites.
Minimum complexity density estimation
Andrew R. Barron and Thomas M. Cover · 1991
Earlier work this paper cites.
The nature of statistical learning theory
Vladimir N. Vapnik · 1995
Earlier work this paper cites.
Probability inequalities for likelihood ratios and convergence rates of sieve MLEs
Wing Hung Wong and Xiaotong Shen · 1995
Earlier work this paper cites.
Rigorous learning curve bounds from statistical mechanics
David Haussler, Michael Kearns, H. Sebastian Seung, and Naftali Tishby · 1996
Earlier work this paper cites.
Efficient agnostic learning of neural networks with bounded fan-in
Wee Sun Lee, Peter L. Bartlett, and Robert C. Williamson · 1996
Earlier work this paper cites.
Mutual information, metric entropy and cumulative relative entropy risk
David Haussler and Manfred Opper · 1997
Earlier work this paper cites.
Minimum contrast estimators on sieves: exponential bounds and rates of convergence
Lucien Birgé and Pascal Massart · 1998
Earlier work this paper cites.
A decision-theoretic extension of stochastic complexity and its applications to learning
Kenji Yamanishi · 1998
Earlier work this paper cites.
An asymptotic property of model selection criteria
Yuhong Yang and Andrew R Barron · 1998
Earlier work this paper cites.
The consistency of posterior distributions in nonparametric problems
Andrew Barron, Mark J. Schervish, and Larry Wasserman · 1999
Earlier work this paper cites.
Viewing all models as “probabilistic”
Peter D. Grünwald · 1999
Earlier work this paper cites.
Estimation of mixture models
Qiang (Jonathan) Li · 1999
Earlier work this paper cites.
Information-theoretic determination of minimax rates of convergence
Yuhong Yang and Andrew Barron · 1999
Earlier work this paper cites.
Convergence rates of posterior distributions
Subhashis Ghosal, Jayanta K. Ghosh, and Aad W. van der Vaart · 2000
Earlier work this paper cites.
Real analysis and probability , volume 74
Richard M. Dudley · 2002
Earlier work this paper cites.
On Bayesian consistency
Stephen Walker and Nils Lid Hjort · 2002
Earlier work this paper cites.
A PAC-Bayesian approach to adaptive classification
Olivier Catoni · 2003
Cited alongside, same era.
PAC-Bayesian stochastic model selection
David McAllester · 2003
Cited alongside, same era.
Generalization error bounds for Bayesian mixture algorithms
R. Meir and T. Zhang · 2003
Cited alongside, same era.
PAC-Bayesian statistical learning theory
Jean-Yves Audibert · 2004
Cited alongside, same era.
Model selection for Gaussian regression with random design
Lucien Birgé · 2004
Cited alongside, same era.
Game theory, maximum entropy, minimum discrepancy and robust Bayesian decision theory
Peter D. Grünwald and A. Philip Dawid · 2004
Cited alongside, same era.
The semiparametric Bernstein–von Mises theorem
Peter J. Bickel and Bas J.K. Kleijn · 2012
Later among the works it cites.
Challenging the empirical mean and empirical variance: a deviation study
Olivier Catoni · 2012
Later among the works it cites.
The safe Bayesian: learning the learning rate via the mixability gap
Peter D. Grünwald · 2012
Later among the works it cites.
Follow the leader if you can, hedge if you must
Steven de Rooij, Tim van Erven, Peter D. Grünwald, and Wouter M. Koolen · 2014
Later among the works it cites.
Optimal learning with Q Q -aggregation
Guillaume Lecué and Philippe Rigollet · 2014
Later among the works it cites.
Learning without concentration
Shahar Mendelson · 2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Alexander B. Tsybakov · 2004
Cited alongside, same era.
Local Rademacher complexities
Peter L. Bartlett, Olivier Bousquet, and Shahar Mendelson · 2005
Cited alongside, same era.
Empirical minimization
Peter L. Bartlett and Shahar Mendelson · 2006
Cited alongside, same era.
Prediction, Learning and Games
Nicòlo Cesa-Bianchi and Gábor Lugosi · 2006
Cited alongside, same era.
Misspecification in infinite-dimensional Bayesian statistics
Bas J.K. Kleijn and Aad W. van der Vaart · 2006
Cited alongside, same era.
Local Rademacher complexities and oracle inequalities in risk minimization
Vladimir Koltchinskii · 2006
Cited alongside, same era.
Rényi divergence and Kullback-Leibler divergence
Tim van Erven and Peter Harremoës · 2014
Later among the works it cites.
Learning with square loss: localization through offset Rademacher complexity
Tengyuan Liang, Alexander Rakhlin, and Karthik Sridharan · 2015
Later among the works it cites.
Fast rates in statistical and online learning
Tim van Erven, Peter D. Grünwald, Nishant A. Mehta, Mark D. Reid, and Robert C. Williamson · 2015
Later among the works it cites.
A general framework for updating belief distributions
Pier Giovanni Bissiri, Chris C. Holmes, and Stephen G. Walker · 2016
Closest in time.
Fast learning rates with heavy-tailed losses
Vu C. Dinh, Lam S. Ho, Binh Nguyen, and Duy Nguyen · 2016
Closest in time.
Loss minimization and parameter estimation with heavy tails
Daniel J. Hsu and Sivan Sabato · 2016
Closest in time.
Combining adversarial guarantees and stochastic fast rates in online learning
Wouter M. Koolen, Peter Grünwald, and Tim van Erven · 2016
Closest in time.
f f -divergence inequalities
Igal Sason and Sergio Verdú · 2016
Closest in time.
Inconsistency of Bayesian inference for misspecified linear models, and a proposal for repairing it
Peter D. Grünwald and Thijs Van Ommen · 2017
Closest in time.
Empirical Bayes posterior concentration in sparse high-dimensional linear models
Ryan Martin, Raymond Mess, and Stephen G. Walker · 2017
Closest in time.
Robust Bayesian inference via coarsening
Jeffrey W Miller and David B Dunson · 2018
Closest in time.
Bayesian fractional posteriors
Anirban Bhattacharya, Debdeep Pati, Yun Yang, et al · 2019
Closest in time.
Relative deviation learning bounds and generalization with unbounded loss functions
Corinna Cortes, Spencer Greenberg, and Mehryar Mohri · 2019
Closest in time.
A tight excess risk bound via a unified PAC-Bayesian-Rademacher-Shtarkov-MDL complexity
Peter D. Grünwald and Nishant A. Mehta · 2019
Closest in time.
Safe-Bayesian generalized linear regression
R. De Heide, A. Kirichenko, P. Grünwald, and N. Mehta · 2019
Closest in time.
Regularization, sparse recovery, and median-of-means tournaments
Gábor Lugosi and Shahar Mendelson · 2019
Closest in time.