Fetching the paper…
Reading the bibliography…
We present TabPFN, a trained Transformer that can do supervised classification for small tabular datasets in less than a second, needs no hyperparameter tuning and is competitive with state-of-the-art classification methods.
Bayesian Learning for Neural Networks
R. Neal · 1996
Earlier work this paper cites.
A new family of power transformations to improve normality or symmetry
I. Yeo and R. Johnson · 2000
Earlier work this paper cites.
Random forests
L. Breimann · 2001
Earlier work this paper cites.
Greedy function approximation: a gradient boosting machine
J. Friedman · 2001
Earlier work this paper cites.
The speed prior: A new simplicity measure yielding near-optimal computable predictions
J. Schmidhuber · 2002
Earlier work this paper cites.
Feature selection, l1 vs. l2 regularization, and rotational invariance
A. Ng · 2004
Earlier work this paper cites.
Statistical comparisons of classifiers over multiple data sets
J. Demšar · 2006
Earlier work this paper cites.
Causality
J. Pearl · 2009
Earlier work this paper cites.
Rectified linear units improve restricted Boltzmann machines
V. Nair and G. Hinton · 2010
Earlier work this paper cites.
Causal inference
J. Pearl · 2010
Earlier work this paper cites.
Scikit-learn: Machine learning in Python
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay · 2011
Earlier work this paper cites.
On causal and anticausal learning
B. Schölkopf, D. Janzing, J. Peters, E. Sgouritsa, K. Zhang, and J. Mooij · 2012
Earlier work this paper cites.
Causal reasoning
M. Waldmann and Y. Hagmayer · 2013
Earlier work this paper cites.
Dropout: A simple way to prevent neural networks from overfitting
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov · 2014
Earlier work this paper cites.
OpenML: Networked science in machine learning
J. Vanschoren, J. van Rijn, B. Bischl, and L. Torgo · 2014
Earlier work this paper cites.
Efficient and robust automated machine learning
M. Feurer, A. Klein, K. Eggensperger, J. Springenberg, M. Blum, and F. Hutter · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. Kingma and J. Ba · 2015
Earlier work this paper cites.
Xgboost: A scalable tree boosting system
T. Chen and C. Guestrin · 2016
Earlier work this paper cites.
Uncertainty in Deep Learning
Y. Gal · 2016
Earlier work this paper cites.
UCI machine learning repository, 2017
Dheeru Dua and Casey Graff · 2017
Cited alongside, same era.
Lightgbm: A highly efficient gradient boosting decision tree
G. Ke, Q. Meng, T. Finley, T. Wang, W. Chen, W. Ma, Q. Ye, and T.-Y. Liu · 2017
Cited alongside, same era.
SGDR: Stochastic gradient descent with warm restarts
I. Loshchilov and F. Hutter · 2017
Cited alongside, same era.
Elements of causal inference: foundations and learning algorithms
J. Peters, D. Janzing, and B. Schölkopf · 2017
Cited alongside, same era.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. Gomez, L. Kaiser, and I. Polosukhin · 2017
Cited alongside, same era.
Neural Architecture Search with reinforcement learning
B. Zoph and Q. V. Le · 2017
Cited alongside, same era.
From probability to consilience: How explanatory values implement bayesian reasoning
Z. Wojtowicz and S. DeDeo · 2020
Later among the works it cites.
Big bird: Transformers for longer sequences
M. Zaheer, G. Guruganesh, K. Dubey, J. Ainslie, C. Alberti, S. Ontanon, P. Pham, A. Ravula, Q. Wang Qifan, L. Yang, and A. Amr · 2020
Later among the works it cites.
TabNet: Attentive interpretable tabular learning
S. Arik and T. Pfister · 2021
Later among the works it cites.
OpenML benchmarking suites
B. Bischl, G. Casalicchio, M. Feurer, P. Gijsbers, F. Hutter, M. Lang, R. Mantovani, J. van Rijn, and J. Vanschoren · 2021
Later among the works it cites.
Deep neural networks and tabular data: A survey
V. Borisov, T. Leemann, K. Seßler, J. Haug, M. Pawelczyk, and G. Kasneci · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Notes from the AI frontier: insights from hundreds of use cases, 2018
M. Chui, J. Manyika, M. Miremadi, N. Henke, R. Chung, P. Nel, and S. Malhotra · 2018
Cited alongside, same era.
The Book of Why
J. Pearl and D. Mackenzie · 2018
Cited alongside, same era.
CatBoost: unbiased boosting with categorical features
L. Prokhorenkova, G. Gusev, A. Vorobev, A. Dorogush, and A. Gulin · 2018
Cited alongside, same era.
Anchor regression: heterogeneous data meets causality
D. Rothenhäusler, N. Meinshausen, P. Bühlmann, and J. Peters · 2018
Cited alongside, same era.
Hyperparameter Optimization
M. Feurer and F. Hutter · 2019
Cited alongside, same era.
Longformer: The long-document transformer
I. Beltagy, M. Peters, and A. Cohan · 2020
Cited alongside, same era.
M. Feurer, K. Eggensperger, S. Falkner, M. Lindauer, and F. Hutter · 2021
Later among the works it cites.
Well-tuned simple nets excel on tabular datasets
A. Kadra, M. Lindauer, F. Hutter, and J. Grabocka · 2021
Later among the works it cites.
Self-attention between datapoints: Going beyond individual input-output pairs in deep learning
J. Kossen, N. Band, C. Lyle, A. Gomez, T. Rainforth, and Y. Gal · 2021
Later among the works it cites.
Miracle: Causally-aware imputation via learning missing data mechanisms
T. Kyono, Y. Zhang, A. Bellot, and M. van der Schaar · 2021
Later among the works it cites.
A scoping review of causal methods enabling predictions under hypothetical interventions
L. Lin, M. Sperrin, D. A Jenkins, G. Martin, and N. Peek · 2021
Later among the works it cites.
SAINT: Improved neural networks for tabular data via row attention and contrastive pre-training
G. Somepalli, M. Goldblum, A. Schwarzschild, C. Bruss, and T. Goldstein · 2021
Later among the works it cites.
Boosting ensemble accuracy by revisiting ensemble diversity metrics
Yanzhao Wu, Ling Liu, Zhongwei Xie, Ka-Ho Chow, and Wenqi Wei · 2021
Later among the works it cites.
Neural ensemble search for uncertainty estimation and dataset shift
S. Zaidi, A. Zela, T. Elsken, C. Holmes, F. Hutter, and Y. Teh · 2021
Later among the works it cites.
Why do tree-based models still outperform deep learning on tabular data?
L. Grinsztajn, E. Oyallon, and G. Varoquaux · 2022
Closest in time.
Learning to induce causal structure, 2022
Nan Rosemary Ke, Silvia Chiappa, Jane Wang, Anirudh Goyal, Jorg Bornschein, Melanie Rey, Theophane Weber, Matthew Botvinic, Michael Mozer, and Danilo Jimenez Rezende · 2022
Closest in time.
Amortized inference for causal structure learning, 2022
Lars Lorch, Scott Sussex, Jonas Rothfuss, Andreas Krause, and Bernhard Schölkopf · 2022
Closest in time.
Transformers can do bayesian inference
S. Müller, N. Hollmann, S. Arango, J. Grabocka, and F. Hutter · 2022
Closest in time.
Tabular data: Deep learning is not all you need
R. Shwartz-Ziv and A. Armon · 2022
Closest in time.
Xgboost search space tweet
Bojan Tunguz · 2022
Closest in time.