Fetching the paper…
Reading the bibliography…
We consider the use of automated supervised learning systems for data tables that not only contain numeric/categorical columns, but one or more text fields as well.
Stacked generalization
D. H. Wolpert · 1992
Earlier work this paper cites.
Random forests
L. Breiman · 2001
Earlier work this paper cites.
Ensemble selection from libraries of models
R. Caruana, A. Niculescu-Mizil, G. Crew, and A. Ksikes · 2004
Earlier work this paper cites.
Ensemble selection from libraries of models
R. Caruana, A. Niculescu-Mizil, G. Crew, and A. Ksikes · 2004
Earlier work this paper cites.
Trade: Transformers for density estimation
R. Fakoor, P. Chaudhari, J. Mueller, and A. J. Smola · 2004
Earlier work this paper cites.
Extremely randomized trees
P. Geurts, D. Ernst, and L. Wehenkel · 2006
Earlier work this paper cites.
Extremely randomized trees
P. Geurts, D. Ernst, and L. Wehenkel · 2006
Earlier work this paper cites.
UCI machine learning repository, 2007
A. Asuncion and D. Newman · 2007
Earlier work this paper cites.
Super learner
M. J. Van der Laan, E. C. Polley, and A. E. Hubbard · 2007
Earlier work this paper cites.
Super learner
M. J. Van der Laan, E. C. Polley, and A. E. Hubbard · 2007
Earlier work this paper cites.
Openml: Networked science in machine learning
J. Vanschoren, J. N. van Rijn, B. Bischl, and L. Torgo · 2013
Earlier work this paper cites.
Efficient estimation of word representations in vector space
T. Mikolov, K. Chen, G. Corrado, and J. Dean · 2013
Earlier work this paper cites.
A proactive intelligent decision support system for predicting the popularity of online news
K. Fernandes, P. Vinagre, and P. Cortez · 2015
Earlier work this paper cites.
Efficient and robust automated machine learning
M. Feurer, A. Klein, K. Eggensperger, J. Springenberg, M. Blum, and F. Hutter · 2015
Earlier work this paper cites.
Taking the human out of the loop: A review of bayesian optimization
B. Shahriari, K. Swersky, Z. Wang, R. P. Adams, and N. De Freitas · 2015
Earlier work this paper cites.
XGBoost: A scalable tree boosting system
T. Chen and C. Guestrin · 2016
Earlier work this paper cites.
Entity embeddings of categorical variables
C. Guo and F. Berkhahn · 2016
Earlier work this paper cites.
Deep crossing: Web-scale modeling without manually crafted combinatorial features
Y. Shan, T. R. Hoens, J. Jiao, H. Wang, D. Yu, and J. Mao · 2016
Earlier work this paper cites.
Entity embeddings of categorical variables
C. Guo and F. Berkhahn · 2016
Earlier work this paper cites.
Deep crossing: Web-scale modeling without manually crafted combinatorial features
Y. Shan, T. R. Hoens, J. Jiao, H. Wang, D. Yu, and J. Mao · 2016
Earlier work this paper cites.
NLP with H2O
H2O.ai · 2017
Earlier work this paper cites.
LightGBM: A highly efficient gradient boosting decision tree
G. Ke, Q. Meng, T. Finley, T. Wang, W. Chen, W. Ma, Q. Ye, and T.-Y. Liu · 2017
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
NLP with H2O
H2O.ai · 2017
Earlier work this paper cites.
LightGBM: A highly efficient gradient boosting decision tree
G. Ke, Q. Meng, T. Finley, T. Wang, W. Chen, W. Ma, Q. Ye, and T.-Y. Liu · 2017
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
Automatic detection of online recruitment frauds: Characteristics, methods, and a public dataset
S. Vidros, C. Kolias, G. Kambourakis, and L. Akoglu · 2017
Earlier work this paper cites.
Data science trends on kaggle
S. Bansal · 2018
Earlier work this paper cites.
Natural language processing, 2018
J. Eisenstein · 2018
Earlier work this paper cites.
Enhanced text classification and word vectors using amazon sagemaker blazingtext
S. Gupta and V. Khare · 2018
Earlier work this paper cites.
Automated Machine Learning: Methods, Systems, Challenges
F. Hutter, L. Kotthoff, and J. Vanschoren · 2018
Earlier work this paper cites.
1st place solution to mercari price suggestion challenge
K. Lopuhin and P. Jankiewicz · 2018
Earlier work this paper cites.
CatBoost: unbiased boosting with categorical features
L. Prokhorenkova, G. Gusev, A. Vorobev, A. V. Dorogush, and A. Gulin · 2018
Earlier work this paper cites.
A. F. Agarap · 2018
Cited alongside, same era.
Natural language processing, 2018
J. Eisenstein · 2018
Cited alongside, same era.
T. Gebru, J. Morgenstern, B. Vecchione, J. W. Vaughan, H. Wallach, H. Daumé III, and K. Crawford · 2018
Cited alongside, same era.
Universal language model fine-tuning for text classification
J. Howard and S. Ruder · 2018
Cited alongside, same era.
1st place solution to mercari price suggestion challenge
K. Lopuhin and P. Jankiewicz · 2018
Cited alongside, same era.
CatBoost: unbiased boosting with categorical features
Adversarial NLI: A new benchmark for natural language understanding
Y. Nie, A. Williams, E. Dinan, M. Bansal, J. Weston, and D. Kiela · 2020
Later among the works it cites.
Pre-trained models for natural language processing: A survey
X. Qiu, T. Sun, Y. Xu, Y. Shao, N. Dai, and X. Huang · 2020
Later among the works it cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu · 2020
Later among the works it cites.
pytorch-widedeep, deep learning for tabular data i: data preprocessing, model components and basic use
J. Rodriguez · 2020
Later among the works it cites.
Mmf: A multimodal framework for vision and language research
A. Singh, V. Goswami, V. Natarajan, Y. Jiang, X. Chen, M. Shah, M. Rohrbach, D. Batra, and D. Parikh · 2020
Later among the works it cites.
VD-BERT: A unified vision and dialog transformer with BERT
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
L. Prokhorenkova, G. Gusev, A. Vorobev, A. V. Dorogush, and A. Gulin · 2018
Cited alongside, same era.
How algorithms discriminate based on data they lack: Challenges, solutions, and policy implications
B. A. Williams, C. F. Brooks, and Y. Shmargad · 2018
Cited alongside, same era.
Multimodal deep networks for text and image-based document classification
N. Audebert, C. Herold, K. Slimani, and C. Vidal · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova · 2019
Cited alongside, same era.
An open source AutoML benchmark
P. Gijsbers, E. LeDell, J. Thomas, S. Poirier, B. Bischl, and J. Vanschoren · 2019
Cited alongside, same era.
Auto-keras: An efficient neural architecture search system
H. Jin, Q. Song, and X. Hu · 2019
Cited alongside, same era.
Deepgbm: A deep learning framework distilled by gbdt for online prediction tasks
G. Ke, Z. Xu, J. Zhang, J. Bian, and T.-Y. Liu · 2019
Cited alongside, same era.
Y. Wang, S. Joty, M. R. Lyu, I. King, C. Xiong, and S. C. Hoi · 2020
Later among the works it cites.
Tabert: Pretraining for joint understanding of textual and tabular data
P. Yin, G. Neubig, W.-t. Yih, and S. Riedel · 2020
Later among the works it cites.
Ensemble squared: A meta automl system
J. Yoo, T. Joseph, D. Yung, S. A. Nasseri, and F. Wood · 2020
Later among the works it cites.
M. Blohm, M. Hanussek, and M. Kintz · 2020
Later among the works it cites.
Extracting training data from large language models
N. Carlini, F. Tramer, E. Wallace, M. Jagielski, A. Herbert-Voss, K. Lee, A. Roberts, T. Brown, D. Song, U. Erlingsson, et al · 2020
Later among the works it cites.
ELECTRA: Pre-training text encoders as discriminators rather than generators
K. Clark, M.-T. Luong, Q. V. Le, and C. D. Manning · 2020
Later among the works it cites.
Autogluon-tabular: Robust and accurate automl for structured data
N. Erickson, J. Mueller, A. Shirkov, H. Zhang, P. Larroy, M. Li, and A. Smola · 2020
Later among the works it cites.
Tabtransformer: Tabular data modeling using contextual embeddings
X. Huang, A. Khetan, M. Cvitkovic, and Z. Karnin · 2020
Later among the works it cites.
Low resource neural machine translation: A benchmark for five african languages
S. M. Lakew, M. Negri, and M. Turchi · 2020
Later among the works it cites.
ALBERT: A lite bert for self-supervised learning of language representations
Z. Lan, M. Chen, S. Goodman, K. Gimpel, P. Sharma, and R. Soricut · 2020
Later among the works it cites.
H2o automl: Scalable automatic machine learning
E. LeDell and S. Poirier · 2020
Later among the works it cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu · 2020
Later among the works it cites.
pytorch-widedeep, deep learning for tabular data i: data preprocessing, model components and basic use
J. Rodriguez · 2020
Later among the works it cites.
Model-based asynchronous hyperparameter optimization
L. C. Tiao, A. Klein, C. Archambeau, and M. Seeger · 2020
Later among the works it cites.
An empirical study on robustness to spurious correlations using pre-trained language models
L. Tu, G. Lalwani, S. Gella, and H. He · 2020
Later among the works it cites.
A neophyte with automl: Evaluating the promises of automatic machine learning tools
O. Bezrukavnikov and R. Linder · 2021
Closest in time.
Features and capabilities of automl natural language
G. Cloud · 2021
Closest in time.
Which machine learning classifiers are best for small datasets? an empirical study
S. Feldman · 2021
Closest in time.
The GEM benchmark: Natural language generation, its evaluation and metrics
S. Gehrmann, T. Adewumi, K. Aggarwal, P. S. Ammanamanchi, A. Anuoluwapo, A. Bosselut, K. R. Chandu, M. Clinciu, D. Das, K. D. Dhole, et al · 2021
Closest in time.
H2O, Fast Scalable Machine Learning, for python , January 2021
H2O.ai · 2021
Closest in time.
Automl: A survey of the state-of-the-art
X. He, K. Zhao, and X. Chu · 2021
Closest in time.
UniT: Multimodal multitask learning with a unified transformer
R. Hu and A. Singh · 2021
Closest in time.
Auto training and fast deployment for state-of-the-art nlp models
HuggingFace · 2021
Closest in time.
Learning transferable visual models from natural language supervision
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al · 2021
Closest in time.
NBDT: Neural-backed decision tree
A. Wan, L. Dunlap, D. Ho, J. Yin, S. Lee, H. Jin, S. Petryk, S. A. Bargal, and J. E. Gonzalez · 2021
Closest in time.
Benchmark and survey of automated machine learning frameworks
M.-A. Zöller and M. F. Huber · 2021
Closest in time.
A survey on recent approaches for natural language processing in low-resource scenarios
M. A. Hedderich, L. Lange, H. Adel, J. Strötgen, and D. Klakow · 2021
Closest in time.
NBDT: Neural-backed decision tree
A. Wan, L. Dunlap, D. Ho, J. Yin, S. Lee, H. Jin, S. Petryk, S. A. Bargal, and J. E. Gonzalez · 2021
Closest in time.