Fetching the paper…
Reading the bibliography…
The recently introduced TabPFN pretrains an In-Context Learning (ICL) transformer on synthetic data to perform tabular data classification.
OpenML: networked science in machine learning
Joaquin Vanschoren, Jan N. Van Rijn, Bernd Bischl, and Luis Torgo · 1931
Earlier work this paper cites.
Gaussian Processes for Regression
Christopher Williams and Carl Rasmussen · 1995
Earlier work this paper cites.
Support vector machines
M.A. Hearst, S.T. Dumais, E. Osuna, J. Platt, and B. Scholkopf · 1998
Earlier work this paper cites.
Predicting clicks: estimating the click-through rate for new ads
Matthew Richardson, Ewa Dominowska, and Robert Ragno · 2007
Earlier work this paper cites.
Scikit-learn: Machine Learning in Python
Fabian Pedregosa, Gaël Varoquaux, Alexandre Gramfort, Vincent Michel, Bertrand Thirion, Olivier Grisel, Mathieu Blondel, Peter Prettenhofer, Ron Weiss, Vincent Dubourg, Jake Vanderplas, Alexandre Passos, David Cournapeau, Matthieu Brucher, Matthieu Perrot, and Édouard Duchesnay · 2011
Earlier work this paper cites.
XGBoost: A Scalable Tree Boosting System
Tianqi Chen and Carlos Guestrin · 2016
Earlier work this paper cites.
LightGBM: A Highly Efficient Gradient Boosting Decision Tree
Guolin Ke, Qi Meng, Thomas Finley, Taifeng Wang, Wei Chen, Weidong Ma, Qiwei Ye, and Tie-Yan Liu · 2017
Earlier work this paper cites.
Machine Learning in Agriculture: A Review
Konstantinos G. Liakos, Patrizia Busato, Dimitrios Moshou, Simon Pearson, and Dionysis Bochtis · 2018
Earlier work this paper cites.
CatBoost: unbiased boosting with categorical features
Liudmila Prokhorenkova, Gleb Gusev, Aleksandr Vorobev, Anna Veronika Dorogush, and Andrey Gulin · 2018
Earlier work this paper cites.
Regularization Learning Networks: Deep Learning for Tabular Datasets
Ira Shavitt and Eran Segal · 2018
Earlier work this paper cites.
XGBoost Model for Chronic Kidney Disease Diagnosis
Adeola Ogunleye and Qing-Guo Wang · 2019
Earlier work this paper cites.
Tabtransformer: Tabular data modeling using contextual embeddings
Xin Huang, Ashish Khetan, Milan Cvitkovic, and Zohar Karnin · 2020
Earlier work this paper cites.
Predicting Hard Rock Pillar Stability Using GBDT, XGBoost, and LightGBM Algorithms
Weizhang Liang, Suizhi Luo, Guoyan Zhao, and Hao Wu · 2020
Earlier work this paper cites.
The Pitfalls of Simplicity Bias in Neural Networks
Harshay Shah, Kaustav Tamuly, Aditi Raghunathan, Prateek Jain, and Praneeth Netrapalli · 2020
Earlier work this paper cites.
VIME: Extending the Success of Self- and Semi-supervised Learning to Tabular Domain
Jinsung Yoon, Yao Zhang, James Jordon, and Mihaela van der Schaar · 2020
Earlier work this paper cites.
Deep Learning Based Recommender System: A Survey and New Perspectives
Shuai Zhang, Lina Yao, Aixin Sun, and Yi Tay · 2020
Earlier work this paper cites.
Tabnet: Attentive interpretable tabular learning
Sercan Ö Arik and Tomas Pfister · 2021
Earlier work this paper cites.
Revisiting Deep Learning Models for Tabular Data
Yury Gorishniy, Ivan Rubachev, Valentin Khrulkov, and Artem Babenko · 2021
Cited alongside, same era.
Well-tuned Simple Nets Excel on Tabular Datasets, November 2021
Arlind Kadra, Marius Lindauer, Frank Hutter, and Josif Grabocka · 2021
Cited alongside, same era.
Net-DNF: Effective Deep Modeling of Tabular Data
Liran Katzir, Gal Elidan, and Ran El-Yaniv · 2021
Cited alongside, same era.
Combining Machine Learning and Computational Chemistry for Predictive Insights Into Chemical Systems
John A. Keith, Valentin Vassilev-Galindo, Bingqing Cheng, Stefan Chmiela, Michael Gastegger, Klaus-Robert Müller, and Alexandre Tkatchenko · 2021
Cited alongside, same era.
Self-Attention Between Datapoints: Going Beyond Individual Input-Output Pairs in Deep Learning
Jannik Kossen, Neil Band, Clare Lyle, Aidan N Gomez, Thomas Rainforth, and Yarin Gal · 2021
In-Context Data Distillation with TabPFN
Junwei Ma, Valentin Thomas, Guangwei Yu, and Anthony Caterini · 2023
Later among the works it cites.
When Do Neural Nets Outperform Boosted Trees on Tabular Data?
Duncan McElfresh, Sujay Khandagale, Jonathan Valverde, Vishak Prasad C, Benjamin Feuer, Chinmay Hegde, Ganesh Ramakrishnan, Micah Goldblum, and Colin White · 2023
Later among the works it cites.
STUNT: Few-shot Tabular Learning with Self-generated Tasks from Unlabeled Tables
Jaehyun Nam, Jihoon Tack, Kyungmin Lee, Hankook Lee, and Jinwoo Shin · 2023
Later among the works it cites.
High dimensional, tabular deep learning with an auxiliary knowledge graph
Camilo Ruiz, Hongyu Ren, Kexin Huang, and Jure Leskovec · 2023
Later among the works it cites.
Self-supervised Representation Learning from Random Data Projectors
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Tabular data: Deep learning is not all you need
Ravid Shwartz-Ziv and Amitai Armon · 2021
Cited alongside, same era.
SAINT: Improved Neural Networks for Tabular Data via Row Attention and Contrastive Pre-Training
Gowthami Somepalli, Micah Goldblum, Avi Schwarzschild, C. Bayan Bruss, and Tom Goldstein · 2021
Cited alongside, same era.
SubTab: Subsetting Features of Tabular Data for Self-Supervised Representation Learning
Talip Ucar, Ehsan Hajiramezanali, and Lindsay Edwards · 2021
Cited alongside, same era.
SCARF: Self-Supervised Contrastive Learning using Random Feature Corruption
Dara Bahri, Heinrich Jiang, Yi Tay, and Donald Metzler · 2022
Cited alongside, same era.
On Embeddings for Numerical Features in Tabular Deep Learning
Yury Gorishniy, Ivan Rubachev, and Artem Babenko · 2022
Cited alongside, same era.
Why do tree-based models still outperform deep learning on tabular data?
Léo Grinsztajn, Edouard Oyallon, and Gaël Varoquaux · 2022
Cited alongside, same era.
Deep Learning for Anomaly Detection: A Review
Guansong Pang, Chunhua Shen, Longbing Cao, and Anton Van Den Hengel · 2022
Cited alongside, same era.
Yi Sui, Tongzi Wu, Jesse C Cresswell, and Ga Wu · 2023
Later among the works it cites.
Towards Foundation Models for Learning on Tabular Data, October 2023
Han Zhang, Xumeng Wen, Shun Zheng, Wei Xu, and Jiang Bian · 2023
Later among the works it cites.
Unlocking the Transferability of Tokens in Deep Models for Tabular Data
Qi-Le Zhou, Han-Jia Ye, Le-Ye Wang, and De-Chuan Zhan · 2023
Later among the works it cites.
XTab: Cross-table Pretraining for Tabular Transformers
Bingzhao Zhu, Xingjian Shi, Nick Erickson, Mu Li, George Karypis, and Mahsa Shoaran · 2023
Later among the works it cites.
A Performance-Driven Benchmark for Feature Selection in Tabular Deep Learning
Valeriia Cherepanova, Roman Levin, Gowthami Somepalli, Jonas Geiping, C Bayan Bruss, Andrew Gordon Wilson, Tom Goldstein, and Micah Goldblum · 2024
Closest in time.
TuneTables: Context Optimization for Scalable Prior-Data Fitted Networks, March 2024
Benjamin Feuer, Robin Tibor Schirrmeister, Valeriia Cherepanova, Chinmay Hegde, Frank Hutter, Micah Goldblum, Niv Cohen, and Colin White · 2024
Closest in time.
CARTE: pretraining and transfer for tabular learning, February 2024
Myung Jun Kim, Léo Grinsztajn, and Gaël Varoquaux · 2024
Closest in time.
What exactly has TabPFN learned to do?, 2024
Calvin McCarter · 2024
Closest in time.
Interpretable Machine Learning for TabPFN, March 2024
David Rundel, Julius Kobialka, Constantin von Crailsheim, Matthias Feurer, Thomas Nagler, and David Rügamer · 2024
Closest in time.
Retrieval & Fine-Tuning for In-Context Tabular Models, June 2024
Valentin Thomas, Junwei Ma, Rasa Hosseinzadeh, Keyvan Golestan, Guangwei Yu, Maksims Volkovs, and Anthony Caterini · 2024
Closest in time.
Making Pre-trained Language Models Great on Tabular Prediction, March 2024
Jiahuan Yan, Bo Zheng, Hongxia Xu, Yiheng Zhu, Danny Z. Chen, Jimeng Sun, Jian Wu, and Jintai Chen · 2024
Closest in time.
Tabular Data: Is Attention All You Need?
Guri Zabërgja, Arlind Kadra, and Josif Grabocka · 2024
Closest in time.