Fetching the paper…
Reading the bibliography…
In this paper, we demonstrate a surprising capability of large language models (LLMs): given only input feature names and a description of a prediction task, they are capable of selecting the most predictive features, with performance rivaling the standard tools of data science.
A New Measure of Rank Correlation
M. G. Kendall · 1938
Earlier work this paper cites.
A Learning Algorithm for Boltzmann Machines
David H. Ackley, Geoffrey E. Hinton, and Terrence J. Sejnowski · 1985
Earlier work this paper cites.
Using the ADAP Learning Algorithm to Forecast the Onset of Diabetes Mellitus
Jack W. Smith, James E. Everhart, William C. Dickson, William C. Knowler, and Richard S. Johannes · 1988
Earlier work this paper cites.
Feature Selection and Feature Extraction for Text Categorization
David D. Lewis · 1992
Earlier work this paper cites.
Comparative Study of Techniques for Large-scale Feature Selection
F.J. Ferri, P. Pudil, M. Hatef, and J. Kittler · 1994
Earlier work this paper cites.
Statlog (German Credit Data)
Hans Hofmann · 1994
Earlier work this paper cites.
Regression Shrinkage and Selection via the Lasso
Robert Tibshirani · 1996
Earlier work this paper cites.
Selection of Relevant Features and Examples in Machine Learning
Avrim L. Blum and Pat Langley · 1997
Earlier work this paper cites.
Wrappers for Feature Subset Selection
Ron Kohavi and George H. John · 1997
Earlier work this paper cites.
Sparse Spatial Autoregressions
R. Kelley Pace and Ronald Barry · 1997
Earlier work this paper cites.
Algorithm 778: L-BFGS-B: Fortran Subroutines for Large-Scale Bound-Constrained Optimization
Ciyou Zhu, Richard H. Byrd, Peihuang Lu, and Jorge Nocedal · 1997
Earlier work this paper cites.
PhysioBank, PhysioToolkit, and PhysioNet: Components of a New Research Resource for Complex Physiologic Signals
Ary L. Goldberger, Luis A. N. Amaral, Leon Glass, Jeffrey M. Hausdorff, Plamen Ch. Ivanov, Roger G. Mark, Joseph E. Mietus, George B. Moody, Chung-Kang Peng, and H. Eugene Stanley · 2000
Earlier work this paper cites.
Random Forests
Leo Breiman · 2001
Earlier work this paper cites.
Pattern Classification
R.O. Duda, P.E. Hart, and D.G. Stork · 2001
Earlier work this paper cites.
Gene Selection for Cancer Classification Using Support Vector Machines
Isabelle Guyon, Jason Weston, Stephen Barnhill, and Vladimir Vapnik · 2002
Earlier work this paper cites.
An Introduction to Variable and Feature Selection
Isabelle Guyon and André Elisseeff · 2003
Earlier work this paper cites.
Least Angle Regression
Bradley Efron, Trevor Hastie, Iain Johnstone, and Robert Tibshirani · 2004
Earlier work this paper cites.
Estimating Mutual Information
Alexander Kraskov, Harald Stögbauer, and Peter Grassberger · 2004
Earlier work this paper cites.
Minimum Redundancy Feature Selection From Microarray Gene Expression Data
Chris Ding and Hanchuan Peng · 2005
Earlier work this paper cites.
Model Selection and Estimation in Regression with Grouped Variables
Ming Yuan and Yi Lin · 2006
Earlier work this paper cites.
Wine Quality
Paulo Cortez, A. Cerdeira, F. Almeida, T. Matos, and J. Reis · 2009
Earlier work this paper cites.
Regularization Paths for Generalized Linear Models via Coordinate Descent
Jerome Friedman, Robert Tibshirani, and Trevor Hastie · 2010
Earlier work this paper cites.
Generalized Fisher Score for Feature Selection
Quanquan Gu, Zhenhui Li, and Jiawei Han · 2011
Earlier work this paper cites.
A Survey on Filter Techniques for Feature Selection in Gene Expression Microarray Analysis
Cosmin Lazar, Jonatan Taminau, Stijn Meganck, David Steenhoff, Alain Coletta, Colin Molter, Virginie de Schaetzen, Robin Duque, Hugues Bersini, and Ann Nowe · 2012
Earlier work this paper cites.
Bank Marketing
S. Moro, P. Rita, and P. Cortez · 2012
Earlier work this paper cites.
Feature Selection via Dependence Maximization
Le Song, Alex Smola, Arthur Gretton, Justin Bedo, and Karsten Borgwardt · 2012
Earlier work this paper cites.
Learning Fair Representations
Rich Zemel, Yu Wu, Kevin Swersky, Toni Pitassi, and Cynthia Dwork · 2013
Earlier work this paper cites.
A Survey on Feature Selection Methods
Girish Chandrashekar and Ferat Sahin · 2014
Earlier work this paper cites.
SAGA: A Fast Incremental Gradient Method With Support for Non-Strongly Convex Composite Objectives
Aaron Defazio, Francis Bach, and Simon Lacoste-Julien · 2014
Earlier work this paper cites.
Sequential Lasso Cum EBIC for Feature Selection With Ultra-High Dimensional Feature Space
Shan Luo and Zehua Chen · 2014
Earlier work this paper cites.
Mutual Information Between Discrete and Continuous Data Sets
Brian C. Ross · 2014
Cited alongside, same era.
High-Dimensional Feature Selection by Feature-Wise Kernelized Lasso
Makoto Yamada, Wittawat Jitkrittum, Leonid Sigal, Eric P. Xing, and Masashi Sugiyama · 2014
Cited alongside, same era.
Feature Selection using Joint Mutual Information Maximisation
Mohamed Bennasar, Yulia Hicks, and Rossitza Setchi · 2015
Cited alongside, same era.
Adam: A Method for Stochastic Optimization
Diederik P. Kingma and Jimmy Ba · 2015
Cited alongside, same era.
XGBoost: A Scalable Tree Boosting System
Tianqi Chen and Carlos Guestrin · 2016
Cited alongside, same era.
How We Analyzed the COMPAS Recidivism Algorithm
Jeff Larson, Julia Angwin, Lauren Kirchner, and Surya Mattu · 2016
Cited alongside, same era.
LMPriors: Pre-Trained Language Models as Task-Specific Priors
Kristy Choi, Chris Cundy, Sanjari Srivastava, and Stefano Ermon · 2022
Later among the works it cites.
PaLM: Scaling Language Modeling with Pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, Parker Schuh, Kensen Shi, Sasha Tsvyashchenko, Joshua Maynez, Abhishek Rao, Parker Barnes, Yi Tay, Noam Shazeer, Vinodkumar Prabhakaran, Emily Reif, Nan Du, Ben Hutchinson, Reiner Pope, James Bradbury, Jacob Austin, Michael Isard, Guy Gur-Ari, Pengcheng Yin, Toju Duke, Anselm Levskaya, Sanjay Ghemawat, Sunipa Dev, Henryk Michalewski, Xavier Garcia, Vedant Misra, Kevin Robinson, Liam Fedus, Denny Zhou, Daphne Ippolito, David Luan, Hyeontaek Lim, Barret Zoph, Alexander Spiridonov, Ryan Sepassi, David Dohan, Shivani Agrawal, Mark Omernick, Andrew M. Dai, Thanumalayan Sankaranarayana Pillai, Marie Pellat, Aitor Lewkowycz, Erica Moreira, Rewon Child, Oleksandr Polozov, Katherine Lee, Zongwei Zhou, Xuezhi Wang, Brennan Saeta, Mark Diaz, Orhan Firat, Michele Catasta, Jason Wei, Kathy Meier-Hellstern, Douglas Eck, Jeff Dean, Slav Petrov, and Noah Fiedel · 2022
Later among the works it cites.
Why Do Tree-based Models Still Outperform Deep Learning on Typical Tabular Data?
Leo Grinsztajn, Edouard Oyallon, and Gael Varoquaux · 2022
Later among the works it cites.
On Feature Learning in the Presence of Spurious Correlations
Pavel Izmailov, Polina Kirichenko, Nate Gruver, and Andrew G Wilson · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kernel Feature Selection via Conditional Covariance Minimization
Jianbo Chen, Mitchell Stern, Martin J Wainwright, and Michael I Jordan · 2017
Cited alongside, same era.
Deep Reinforcement Learning from Human Preferences
Paul F Christiano, Jan Leike, Tom Brown, Miljan Martic, Shane Legg, and Dario Amodei · 2017
Cited alongside, same era.
Sparse-input Neural Networks for High-dimensional Nonparametric Regression and Classification
Jean Feng and Noah Simon · 2017
Cited alongside, same era.
LightGBM: A Highly Efficient Gradient Boosting Decision Tree
Guolin Ke, Qi Meng, Thomas Finley, Taifeng Wang, Wei Chen, Weidong Ma, Qiwei Ye, and Tie-Yan Liu · 2017
Cited alongside, same era.
Feature Selection: A Data Perspective
Jundong Li, Kewei Cheng, Suhang Wang, Fred Morstatter, Robert P. Trevino, Jiliang Tang, and Huan Liu · 2017
Cited alongside, same era.
A Unified Approach to Interpreting Model Predictions
Scott M. Lundberg and Su-In Lee · 2017
Cited alongside, same era.
Large Language Models are Zero-Shot Reasoners
Takeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa · 2022
Later among the works it cites.
Solving Quantitative Reasoning Problems with Language Models
Aitor Lewkowycz, Anders Johan Andreassen, David Dohan, Ethan Dyer, Henryk Michalewski, Vinay Venkatesh Ramasesh, Ambrose Slone, Cem Anil, Imanol Schlag, Theo Gutman-Solo, Yuhuai Wu, Behnam Neyshabur, Guy Gur-Ari, and Vedant Misra · 2022
Later among the works it cites.
TruthfulQA: Measuring How Models Mimic Human Falsehoods
Stephanie Lin, Jacob Hilton, and Owain Evans · 2022
Later among the works it cites.
Can Large Language Models Reason About Medical Questions?
Valentin Liévin, Christoffer Egeberg Hother, and Ole Winther · 2022
Later among the works it cites.
Training Language Models to Follow Instructions with Human Feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul F Christiano, Jan Leike, and Ryan Lowe · 2022
Later among the works it cites.
Mapping Language Models to Grounded Conceptual Spaces
Roma Patel and Ellie Pavlick · 2022
Later among the works it cites.
Learning to Summarize from Human Feedback
Nisan Stiennon, Long Ouyang, Jeff Wu, Daniel M. Ziegler, Ryan Lowe, Chelsea Voss, Alec Radford, Dario Amodei, and Paul Christiano · 2022
Later among the works it cites.
PaLM 2 Technical Report, 2023
Rohan Anil, Andrew M. Dai, Orhan Firat, Melvin Johnson, Dmitry Lepikhin, Alexandre Passos, Siamak Shakeri, Emanuel Taropa, Paige Bailey, Zhifeng Chen, Eric Chu, Jonathan H. Clark, Laurent El Shafey, Yanping Huang, Kathy Meier-Hellstern, Gaurav Mishra, Erica Moreira, Mark Omernick, Kevin Robinson, Sebastian Ruder, Yi Tay, Kefan Xiao, Yuanzhong Xu, Yujing Zhang, Gustavo Hernandez Abrego, Junwhan Ahn, Jacob Austin, Paul Barham, Jan Botha, James Bradbury, Siddhartha Brahma, Kevin Brooks, Michele Catasta, Yong Cheng, Colin Cherry, Christopher A. Choquette-Choo, Aakanksha Chowdhery, Clément Crepy, Shachi Dave, Mostafa Dehghani, Sunipa Dev, Jacob Devlin, Mark Díaz, Nan Du, Ethan Dyer, Vlad Feinberg, Fangxiaoyu Feng, Vlad Fienber, Markus Freitag, Xavier Garcia, Sebastian Gehrmann, Lucas Gonzalez, Guy Gur-Ari, Steven Hand, Hadi Hashemi, Le Hou, Joshua Howland, Andrea Hu, Jeffrey Hui, Jeremy Hurwitz, Michael Isard, Abe Ittycheriah, Matthew Jagielski, Wenhao Jia, Kathleen Kenealy, Maxim Krikun, Sneha Kudugunta, Chang Lan, Katherine Lee, Benjamin Lee, Eric Li, Music Li, Wei Li, YaGuang Li, Jian Li, Hyeontaek Lim, Hanzhao Lin, Zhongtao Liu, Frederick Liu, Marcello Maggioni, Aroma Mahendru, Joshua Maynez, Vedant Misra, Maysam Moussalem, Zachary Nado, John Nham, Eric Ni, Andrew Nystrom, Alicia Parrish, Marie Pellat, Martin Polacek, Alex Polozov, Reiner Pope, Siyuan Qiao, Emily Reif, Bryan Richter, Parker Riley, Alex Castro Ros, Aurko Roy, Brennan Saeta, Rajkumar Samuel, Renee Shelby, Ambrose Slone, Daniel Smilkov, David R. So, Daniel Sohn, Simon Tokumine, Dasha Valter, Vijay Vasudevan, Kiran Vodrahalli, Xuezhi Wang, Pidong Wang, Zirui Wang, Tao Wang, John Wieting, Yuhuai Wu, Kelvin Xu, Yunhan Xu, Linting Xue, Pengcheng Yin, Jiahui Yu, Qiao Zhang, Steven Zheng, Ce Zheng, Weikang Zhou, Denny Zhou, Slav Petrov, and Yonghui Wu · 2023
Later among the works it cites.
When Do You Need Chain-of-Thought Prompting for ChatGPT?
Jiuhai Chen, Lichang Chen, Heng Huang, and Tianyi Zhou · 2023
Later among the works it cites.
MathPrompter: Mathematical Reasoning using Large Language Models
Shima Imani, Liang Du, and Harsh Shrivastava · 2023
Later among the works it cites.
MIMIC-IV, A Freely Accessible Electronic Health Record Dataset
Alistair E. W. Johnson, Lucas Bulgarelli, Lu Shen, Alvin Gayles, Ayad Shammout, Steven Horng, Tom J. Pollard, Sicheng Hao, Benjamin Moody, Brian Gow, Li-wei H. Lehman, Leo A. Celi, and Roger G. Mark · 2023
Later among the works it cites.
Efficient Memory Management for Large Language Model Serving with PagedAttention
Woosuk Kwon, Zhuohan Li, Siyuan Zhuang, Ying Sheng, Lianmin Zheng, Cody Hao Yu, Joseph E. Gonzalez, Hao Zhang, and Ion Stoica · 2023
Later among the works it cites.
Pre-Train, Prompt, and Predict: A Systematic Survey of Prompting Methods in Natural Language Processing
Pengfei Liu, Weizhe Yuan, Jinlan Fu, Zhengbao Jiang, Hiroaki Hayashi, and Graham Neubig · 2023
Later among the works it cites.
Language Models Are Weak Learners
Hariharan Manikandan, Yiding Jiang, and J Zico Kolter · 2023
Later among the works it cites.
Foundation Models for Generalist Medical Artificial Intelligence
Michael Moor, Oishi Banerjee, Zahra Shakeri, Harlan Krumholz, Jure Leskovec, Eric Topol, and Pranav Rajpurkar · 2023
Later among the works it cites.
OpenAI · 2023
Later among the works it cites.
Large Language Models Encode Clinical Knowledge
Karan Singhal, Shekoofeh Azizi, Tao Tu, S. Mahdavi, Jason Wei, Hyung Chung, Nathan Scales, Ajay Tanwani, Heather Cole-Lewis, Stephen Pfohl, Perry Payne, Martin Seneviratne, Paul Gamble, Chris Kelly, Abubakr Babiker, Nathanael Schärli, Aakanksha Chowdhery, Philip Mansfield, Dina Demner-Fushman, and Vivek Natarajan · 2023
Later among the works it cites.
Beyond the Imitation Game: Quantifying and Extrapolating the Capabilities of Language Models
Aarohi Srivastava et al · 2023
Later among the works it cites.
Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them
Mirac Suzgun, Nathan Scales, Nathanael Schärli, Sebastian Gehrmann, Yi Tay, Hyung Won Chung, Aakanksha Chowdhery, Quoc Le, Ed Chi, Denny Zhou, and Jason Wei · 2023
Later among the works it cites.
Llama 2: Open Foundation and Fine-Tuned Chat Models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, Dan Bikel, Lukas Blecher, Cristian Canton Ferrer, Moya Chen, Guillem Cucurull, David Esiobu, Jude Fernandes, Jeremy Fu, Wenyin Fu, Brian Fuller, Cynthia Gao, Vedanuj Goswami, Naman Goyal, Anthony Hartshorn, Saghar Hosseini, Rui Hou, Hakan Inan, Marcin Kardas, Viktor Kerkez, Madian Khabsa, Isabel Kloumann, Artem Korenev, Punit Singh Koura, Marie-Anne Lachaux, Thibaut Lavril, Jenya Lee, Diana Liskovich, Yinghai Lu, Yuning Mao, Xavier Martinet, Todor Mihaylov, Pushkar Mishra, Igor Molybog, Yixin Nie, Andrew Poulton, Jeremy Reizenstein, Rashi Rungta, Kalyan Saladi, Alan Schelten, Ruan Silva, Eric Michael Smith, Ranjan Subramanian, Xiaoqing Ellen Tan, Binh Tang, Ross Taylor, Adina Williams, Jian Xiang Kuan, Puxin Xu, Zheng Yan, Iliyan Zarov, Yuchen Zhang, Angela Fan, Melanie Kambadur, Sharan Narang, Aurelien Rodriguez, Robert Stojnic, Sergey Edunov, and Thomas Scialom · 2023
Later among the works it cites.
Self-Consistency Improves Chain of Thought Reasoning in Language Models
Xuezhi Wang, Jason Wei, Dale Schuurmans, Quoc V Le, Ed H. Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou · 2023
Later among the works it cites.
Tree of Thoughts: Deliberate Problem Solving with Large Language Models
Shunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran, Tom Griffiths, Yuan Cao, and Karthik Narasimhan · 2023
Later among the works it cites.
Sequential Attention for Feature Selection
Taisuke Yasuda, Mohammadhossein Bateni, Lin Chen, Matthew Fahrbach, Gang Fu, and Vahab Mirrokni · 2023
Later among the works it cites.
Graph of Thoughts: Solving Elaborate Problems with Large Language Models
Maciej Besta, Nils Blach, Ales Kubicek, Robert Gerstenberger, Lukas Gianinazzi, Joanna Gajda, Tomasz Lehmann, Michał Podstawski, Hubert Niewiadomski, Piotr Nyczyk, and Torsten Hoefler · 2024
Closest in time.
Bias and Fairness in Large Language Models: A Survey
Isabel O. Gallegos, Ryan A. Rossi, Joe Barrow, Md Mehrab Tanjim, Sungchul Kim, Franck Dernoncourt, Tong Yu, Ruiyi Zhang, and Nesreen K. Ahmed · 2024
Closest in time.
Quantifying Language Models’ Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formatting
Melanie Sclar, Yejin Choi, Yulia Tsvetkov, and Alane Suhr · 2024
Closest in time.