Fetching the paper…
Reading the bibliography…
Deep learning (DL) models for tabular data problems (e.g.
Local learning algorithms
L. Bottou and V. Vapnik · 1992
Earlier work this paper cites.
Scaling up the accuracy of naive-bayes classifiers: a decision-tree hybrid
R. Kohavi · 1996
Earlier work this paper cites.
Sparse spatial autoregressions
R. Kelley Pace and R. Barry · 1997
Earlier work this paper cites.
Comparative accuracies of artificial neural networks and discriminant analysis in predicting forest cover types from cartographic variables
J. A. Blackard and D. J. Dean · 2000
Earlier work this paper cites.
Scikit-learn: Machine learning in Python
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay · 2011
Earlier work this paper cites.
An Introduction to Statistical Learning
G. James, D. Witten, T. Hastie, and R. Tibshirani · 2013
Earlier work this paper cites.
Introducing LETOR 4.0 datasets
T. Qin and T. Liu · 2013
Earlier work this paper cites.
Searching for exotic particles in high-energy physics with deep learning
P. Baldi, P. Sadowski, and D. Whiteson · 2014
Earlier work this paper cites.
Openml: networked science in machine learning
J. Vanschoren, J. N. van Rijn, B. Bischl, and L. Torgo · 2014
Earlier work this paper cites.
Layer normalization
J. L. Ba, J. R. Kiros, and G. E. Hinton · 2016
Earlier work this paper cites.
Xgboost: A scalable tree boosting system
T. Chen and C. Guestrin · 2016
Earlier work this paper cites.
Deep kernel learning
A. G. Wilson, Z. Hu, R. Salakhutdinov, and E. P. Xing · 2016
Earlier work this paper cites.
Lightgbm: A highly efficient gradient boosting decision tree
G. Ke, Q. Meng, T. Finley, T. Wang, W. Chen, W. Ma, Q. Ye, and T.-Y. Liu · 2017
Earlier work this paper cites.
Self-normalizing neural networks
G. Klambauer, T. Unterthiner, A. Mayr, and S. Hochreiter · 2017
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
Gpytorch: Blackbox matrix-matrix gaussian process inference with gpu acceleration
J. R. Gardner, G. Pleiss, D. Bindel, K. Q. Weinberger, and A. G. Wilson · 2018
Earlier work this paper cites.
Catboost: unbiased boosting with categorical features
L. Prokhorenkova, G. Gusev, A. Vorobev, A. V. Dorogush, and A. Gulin · 2018
Earlier work this paper cites.
J. Zhao and K. Cho · 2018
Cited alongside, same era.
Optuna: A next-generation hyperparameter optimization framework
T. Akiba, S. Sano, T. Yanase, T. Ohta, and M. Koyama · 2019
Cited alongside, same era.
Attentive neural processes
H. Kim, A. Mnih, J. Schwarz, M. Garnelo, S. M. A. Eslami, D. Rosenbaum, O. Vinyals, and Y. W. Teh · 2019
Cited alongside, same era.
Decoupled weight decay regularization
I. Loshchilov and F. Hutter · 2019
Cited alongside, same era.
Retrieval augmented language model pre-training
K. Guu, K. Lee, Z. Tung, P. Pasupat, and M. Chang · 2020
Cited alongside, same era.
The tree ensemble layer: Differentiability meets conditional computation
H. Hazimeh, N. Ponomareva, P. Mol, Z. Tan, and R. Mazumder · 2020
Self-attention between datapoints: Going beyond individual input-output pairs in deep learning
J. Kossen, N. Band, C. Lyle, A. N. Gomez, T. Rainforth, and Y. Gal · 2021
Later among the works it cites.
Shifts: A dataset of real distributional shift across multiple large-scale tasks
A. Malinin, N. Band, G. Chesnokov, Y. Gal, M. J. F. Gales, A. Noskov, A. Ploskonosov, L. Prokhorenkova, I. Provilkov, V. Raina, V. Raina, M. Shmatova, P. Tigas, and B. Yangel · 2021
Later among the works it cites.
Retrieval & interaction machine for tabular data prediction
J. Qin, W. Zhang, R. Su, Z. Liu, W. Liu, R. Tang, X. He, and Y. Yu · 2021
Later among the works it cites.
Hopfield networks is all you need
H. Ramsauer, B. Schäfl, J. Lehner, P. Seidl, M. Widrich, L. Gruber, M. Holzleitner, T. Adler, D. P. Kreil, M. K. Kopp, G. Klambauer, J. Brandstetter, and S. Hochreiter · 2021
Later among the works it cites.
SAINT: improved neural networks for tabular data via row attention and contrastive pre-training
G. Somepalli, M. Goldblum, A. Schwarzschild, C. B. Bruss, and T. Goldstein · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Tabtransformer: Tabular data modeling using contextual embeddings
X. Huang, A. Khetan, M. Cvitkovic, and Z. Karnin · 2020
Cited alongside, same era.
Generalization through memorization: Nearest neighbor language models
U. Khandelwal, O. Levy, D. Jurafsky, L. Zettlemoyer, and M. Lewis · 2020
Cited alongside, same era.
Retrieval-augmented generation for knowledge-intensive NLP tasks
P. S. H. Lewis, E. Perez, A. Piktus, F. Petroni, V. Karpukhin, N. Goyal, H. Küttler, M. Lewis, W. Yih, T. Rocktäschel, S. Riedel, and D. Kiela · 2020
Cited alongside, same era.
Neural oblivious decision ensembles for deep learning on tabular data
S. Popov, S. Morozov, and A. Babenko · 2020
Cited alongside, same era.
User behavior retrieval for click-through rate prediction
J. Qin, W. Zhang, X. Wu, J. Jin, Y. Fang, and Y. Yu · 2020
Cited alongside, same era.
Dcn v2: Improved deep & cross network and practical lessons for web-scale learning to rank systems
R. Wang, R. Shivanna, D. Z. Cheng, S. Jain, D. Lin, L. Hong, and E. H. Chi · 2020
Cited alongside, same era.
Later among the works it cites.
Improving language models by retrieving from trillions of tokens
S. Borgeaud, A. Mensch, J. Hoffmann, T. Cai, E. Rutherford, K. Millican, G. van den Driessche, J. Lespiau, B. Damoc, A. Clark, D. de Las Casas, A. Guy, J. Menick, R. Ring, T. Hennigan, S. Huang, L. Maggiore, C. Jones, A. Cassirer, A. Brock, M. Paganini, G. Irving, O. Vinyals, S. Osindero, K. Simonyan, J. W. Rae, E. Elsen, and L. Sifre · 2022
Later among the works it cites.
Learning enhanced representation for tabular data via neighborhood propagation
K. Du, W. Zhang, R. Zhou, Y. Wang, X. Zhao, J. Jin, Q. Gan, Z. Zhang, and D. P. Wipf · 2022
Later among the works it cites.
On embeddings for numerical features in tabular deep learning
Y. Gorishniy, I. Rubachev, and A. Babenko · 2022
Later among the works it cites.
Why do tree-based models still outperform deep learning on typical tabular data?
L. Grinsztajn, E. Oyallon, and G. Varoquaux · 2022
Later among the works it cites.
A memory transformer network for incremental learning
A. Iscen, T. Bird, M. Caron, A. Fathi, and C. Schmid · 2022
Later among the works it cites.
Few-shot learning with retrieval augmented language models
G. Izacard, P. S. H. Lewis, M. Lomeli, L. Hosseini, F. Petroni, T. Schick, J. Dwivedi-Yu, A. Joulin, S. Riedel, and E. Grave · 2022
Later among the works it cites.
Retrieval augmented classification for long-tail visual recognition
A. Long, W. Yin, T. Ajanthan, V. Nguyen, P. Purkait, R. Garg, A. Blair, C. Shen, and A. van den Hengel · 2022
Later among the works it cites.
Dnnr: Differential nearest neighbors regression
Y. Nader, L. Sixt, and T. Landgraf · 2022
Later among the works it cites.
Hopular: Modern hopfield networks for tabular data
B. Schäfl, L. Gruber, A. Bitto-Nemling, and S. Hochreiter · 2022
Later among the works it cites.
S. Wang, Y. Xu, Y. Fang, Y. Liu, S. Sun, R. Xu, C. Zhu, and M. Zeng · 2022
Later among the works it cites.
A flexible nadaraya-watson head can offer explainable and calibrated classification
A. Q. Wang and M. R. Sabuncu · 2023
Closest in time.