Language models are few-shot learners
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., Agarwal, S., Herbert-Voss, A., Krueger, G., Henighan, T., Child, R., Ramesh, A., Ziegler, D., Wu, J., Winter, C., Hesse, C., Chen, M., Sigler, E., Litwin, M., Gray, S., Chess, B., Clark, J., Berner, C., McCandlish, S., Radford, A., Sutskever, I., and Amodei, D. (2020) · 1901
Earlier work this paper cites.
Language models are few-shot learners
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., Agarwal, S., Herbert-Voss, A., Krueger, G., Henighan, T., Child, R., Ramesh, A., Ziegler, D., Wu, J., Winter, C., Hesse, C., Chen, M., Sigler, E., Litwin, M., Gray, S., Chess, B., Clark, J., Berner, C., McCandlish, S., Radford, A., Sutskever, I., and Amodei, D. (2020) · 1901
Earlier work this paper cites.
Using the ADAP Learning Algorithm to Forecast the Onset of Diabetes Mellitus
Smith, J. W., Everhart, J., Dickson, W., Knowler, W., and Johannes, R. (1988) · 1988
Earlier work this paper cites.
International application of a new probability algorithm for the diagnosis of coronary artery disease
Detrano, R., Janosi, A., Steinbrunn, W., Pfisterer, M., Schmid, J.-J., Sandhu, S., Guppy, K. H., Lee, S., and Froelicher, V. (1989) · 1989
Earlier work this paper cites.
Scaling up the accuracy of naive-bayes classifiers: A decision-tree hybrid
Kohavi, R. et al. (1996) · 1996
Earlier work this paper cites.
Sparse spatial autoregressions
Pace, R. K. and Barry, R. (1997) · 1997
Earlier work this paper cites.
Knowledge discovery on rfm model using bernoulli sequence
Yeh, I.-C., Yang, K.-J., and Ting, T.-M. (2009) · 2009
Earlier work this paper cites.
Scikit-learn: Machine learning in Python
Pedregosa, F., Varoquaux, G., Gramfort, A., Michel, V., Thirion, B., Grisel, O., Blondel, M., Prettenhofer, P., Weiss, R., Dubourg, V., Vanderplas, J., Passos, A., Cournapeau, D., Brucher, M., Perrot, M., and Duchesnay, E. (2011) · 2011
Earlier work this paper cites.
TabTransformer: Tabular Data Modeling Using Contextual Embeddings
Original
Huang, X., Khetan, A., Cvitkovic, M., and Karnin, Z. (2020) · 2012
Earlier work this paper cites.
A data-driven approach to predict the success of bank telemarketing
Moro, S., Cortez, P., and Rita, P. (2014) · 2014
Earlier work this paper cites.
Endgame analysis of dou shou qi
van Rijn, J. N. and Vis, J. K. (2014) · 2014
Earlier work this paper cites.
Deep Neural Decision Forests
Kontschieder, P., Fiterau, M., Criminisi, A., and Bulo, S. R. (2015) · 2015
Earlier work this paper cites.
Observational Health Data Sciences and Informatics (OHDSI): Opportunities for Observational Researchers
Hripcsak, G., Duke, J. D., Shah, N. H., Reich, C. G., Huser, V., Schuemie, M. J., Suchard, M. A., Park, R. W., Wong, I. C. K., Rijnbeek, P. R., van der Lei, J., Pratt, N., Noré, N, G. N., Li, Y.-C., Stang, P. E., Madigan, D., and Ryan, P. B. (2015) · 2015
Earlier work this paper cites.
XGBoost: A Scalable Tree Boosting System
Chen, T. and Guestrin, C. (2016) · 2016
Earlier work this paper cites.
Randomized trial of communication facilitators to reduce family distress and intensity of end-of-life care
Curtis, J. R., Treece, P. D., Nielsen, E. L., Gold, J., Ciechanowski, P. S., Shannon, S. E., Khandelwal, N., Young, J. P., and Engelberg, R. A. (2016) · 2016
Earlier work this paper cites.
LightGBM: A Highly Efficient Gradient Boosting Decision Tree
Ke, G., Meng, Q., Finley, T., Wang, T., Chen, W., Ma, W., Ye, Q., and Liu, T.-Y. (2017) · 2017
Earlier work this paper cites.
UCI machine learning repository
Dua, D. and Graff, C. (2017) · 2017
Earlier work this paper cites.
Improving palliative care with deep learning
Avati, A., Jung, K., Harman, S., Downing, L., Ng, A., and Shah, N. H. (2018) · 2018
Earlier work this paper cites.
Improving palliative care with deep learning
Avati, A., Jung, K., Harman, S., Downing, L., Ng, A., and Shah, N. H. (2018) · 2018
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Devlin, J., Chang, M.-W., Lee, K., and Toutanova, K. (2019) · 2019
Earlier work this paper cites.
How many rare diseases are there?
Haendel, M., Vasilevsky, N., Unni, D., Bologa, C., Harris, N., Rehm, H., Hamosh, A., Baynam, G., Groza, T., McMurry, J., et al. (2020) · 2020
Earlier work this paper cites.
Deep entity matching with pre-trained language models
Li, Y., Li, J., Suhara, Y., Doan, A., and Tan, W.-C. (2020) · 2020
Earlier work this paper cites.