Fetching the paper…
Reading the bibliography…
Recent large language models (LLMs) have demonstrated remarkable prediction performance for a growing array of tasks.
Jones, K.S.: A statistical interpretation of term specificity and its application in retrieval. Journal of documentation (1972)
1972
Earlier work this paper cites.
Breiman, L., Friedman, J.H., Olshen, R.A., Stone, C.J.: Classification and Regression Trees. Wadsworth and Brooks, Monterey, CA (1984). https://www.routledge.com/Classification-and-Regression-Trees/Breiman-Friedman-Stone-Olshen/p/book/9780412048418
1984
Earlier work this paper cites.
Hastie, T., Tibshirani, R.: Generalized additive models. Statistical Science 1
1986
Earlier work this paper cites.
Quinlan, J.R.: Induction of decision trees. Machine learning 1
1986
Earlier work this paper cites.
Freund, Y., Schapire, R.E., et al
1996
Earlier work this paper cites.
Breiman, L.: Random forests. Machine Learning 45
2001
Earlier work this paper cites.
Friedman, J.H.: Greedy function approximation: a gradient boosting machine. Annals of statistics, 1189–1232 (2001)
2001
Earlier work this paper cites.
Pang, B., Lee, L.: Seeing stars: Exploiting class relationships for sentiment categorization with respect to rating scales. In: Proceedings of the ACL (2005)
2005
Earlier work this paper cites.
Friedman, J.H., Popescu, B.E.: Predictive learning via rule ensembles. The Annals of Applied Statistics 2
2008
Earlier work this paper cites.
Chipman, H.A., George, E.I., McCulloch, R.E.: Bart: Bayesian additive regression trees. The Annals of Applied Statistics 4
2010
Earlier work this paper cites.
Pedregosa, F., Varoquaux, G.ë.l., Gramfort, A., Michel, V., Thirion, B., Grisel, O., Blondel, M., Prettenhofer, P., Weiss, R., Dubourg, V., et al
2011
Earlier work this paper cites.
Dwork, C., Hardt, M., Pitassi, T., Reingold, O., Zemel, R.: Fairness through awareness. In: Proceedings of the 3rd Innovations in Theoretical Computer Science Conference, pp. 214–226 (2012). ACM
2012
Earlier work this paper cites.
Brennan, T., Oliver, W.L.: The emergence of machine learning techniques in criminology. Criminology & Public Policy 12
2013
Earlier work this paper cites.
Socher, R., Perelygin, A., Wu, J., Chuang, J., Manning, C.D., Ng, A., Potts, C.: Recursive deep models for semantic compositionality over a sentiment treebank. In: Proceedings of the 2013 Conference on Empirical Methods in Natural Language Processing, pp. 1631–1642 (2013)
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
Mikolov, T., Sutskever, I., Chen, K., Corrado, G.S., Dean, J.: Distributed representations of words and phrases and their compositionality. Advances in neural information processing systems 26
2013
Earlier work this paper cites.
Malo, P., Sinha, A., Korhonen, P., Wallenius, J., Takala, P.: Good debt or bad debt: Detecting semantic orientations in economic texts. Journal of the Association for Information Science and Technology 65
2014
Earlier work this paper cites.
Pennington, J., Socher, R., Manning, C.D.: Glove: Global vectors for word representation. In: Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP), pp. 1532–1543 (2014)
2014
Earlier work this paper cites.
Quinlan, J.R.: C4. 5: Programs for Machine Learning. Elsevier, ??? (2014)
2014
Earlier work this paper cites.
Caruana, R., Lou, Y., Gehrke, J., Koch, P., Sturm, M., Elhadad, N.: Intelligible models for healthcare: Predicting pneumonia risk and hospital 30-day readmission. In: Proceedings of the 21th ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, pp. 1721–1730 (2015)
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
Angermueller, C., Pärnamaa, T., Parts, L., Stegle, O.: Deep learning for computational biology. Molecular systems biology 12
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
Ribeiro, M.T., Singh, S., Guestrin, C.: Why should i trust you?: Explaining the predictions of any classifier. In: Proceedings of the 22nd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, pp. 1135–1144 (2016). ACM
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
Huth, A.G., De Heer, W.A., Griffiths, T.L., Theunissen, F.E., Gallant, J.L.: Natural speech reveals the semantic maps that tile human cerebral cortex. Nature 532
2016
Earlier work this paper cites.
2016
Cited alongside, same era.
Chen, T., Guestrin, C.: Xgboost: A scalable tree boosting system. In: Proceedings of the 22nd Acm Sigkdd International Conference on Knowledge Discovery and Data Mining, pp. 785–794 (2016)
2016
Cited alongside, same era.
Honnibal, M., Montani, I.: spaCy 2: Natural language understanding with Bloom embeddings, convolutional neural networks and incremental parsing. To appear (2017)
2017
Cited alongside, same era.
Bertsimas, D., Dunn, J.: Optimal classification trees. Machine Learning 106
2017
Cited alongside, same era.
Koh, P.W., Nguyen, T., Tang, Y.S., Mussmann, S., Pierson, E., Kim, B., Liang, P.: Concept bottleneck models. In: International Conference on Machine Learning, pp. 5338–5348 (2020). PMLR
2020
Later among the works it cites.
Singh, C., Ha, W., Lanusse, F., Boehm, V., Liu, J., Yu, B.: Transformation Importance with Applications to Cosmology (2020)
2020
Later among the works it cites.
2020
Later among the works it cites.
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2018
Cited alongside, same era.
Saravia, E., Liu, H.-C.T., Huang, Y.-H., Wu, J., Chen, Y.-S.: Carer: Contextualized affect representations for emotion recognition. In: Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing, pp. 3687–3697 (2018)
2018
Cited alongside, same era.
Peters, M.E., Neumann, M., Iyyer, M., Gardner, M., Clark, C., Lee, K., Zettlemoyer, L.: Deep contextualized word representations. In: Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers), pp. 2227–2237. Association for Computational Linguistics, New Orleans, Louisiana (2018). https://doi.org/10.18653/v1/N18-1202 . https://aclanthology.org/N18-1202
2018
Cited alongside, same era.
Carreira-Perpinán, M.A., Tavallali, P.: Alternating optimization of decision trees, with application to learning sparse oblique trees. Advances in neural information processing systems 31
2018
Cited alongside, same era.
Li, O., Liu, H., Chen, C., Rudin, C.: Deep learning for case-based reasoning through prototypes: A neural network that explains its predictions. In: Proceedings of the AAAI Conference on Artificial Intelligence, vol. 32 (2018)
2018
Cited alongside, same era.
2018
Cited alongside, same era.
Ha, W., Singh, C., Lanusse, F., Upadhyayula, S., Yu, B.: Adaptive wavelet distillation from neural networks through interpretations. Advances in Neural Information Processing Systems 34
2021
Later among the works it cites.
Wang, B., Komatsuzaki, A.: GPT-J-6B: A 6 Billion Parameter Autoregressive Language Model. https://github.com/kingoflolz/mesh-transformer-jax (2021)
2021
Later among the works it cites.
Schrimpf, M., Blank, I.A., Tuckute, G., Kauf, C., Hosseini, E.A., Kanwisher, N., Tenenbaum, J.B., Fedorenko, E.: The neural architecture of language: Integrative modeling converges on predictive processing. Proceedings of the National Academy of Sciences 118
2021
Later among the works it cites.
Agarwal, R., Melnick, L., Frosst, N., Zhang, X., Lengerich, B., Caruana, R., Hinton, G.E.: Neural additive models: Interpretable machine learning with neural nets. Advances in Neural Information Processing Systems 34
2021
Later among the works it cites.
Singh, C., Nasseri, K., Tan, Y.S., Tang, T., Yu, B.: imodels: a python package for fitting interpretable models. Journal of Open Source Software 6
2021
Later among the works it cites.
Janizek, J.D., Sturmfels, P., Lee, S.-I.: Explaining explanations: Axiomatic feature interactions for deep networks. J. Mach. Learn. Res. 22
2021
Later among the works it cites.
2021
Later among the works it cites.
Kornblith, A.E., Singh, C., Devlin, G., Addo, N., Streck, C.J., Holmes, J.F., Kuppermann, N., Grupp-Phelan, J., Fineman, J., Butte, A.J., Yu, B.: Predictability and stability testing to assess clinical decision instrument performance for children after blunt torso trauma. PLOS Digital Health (2022) https://doi.org/10.1371/journal.pdig.0000076
2022
Closest in time.
2022
Closest in time.
LeBel, A., Wagner, L., Jain, S., Adhikari-Desai, A., Gupta, B., Morgenthal, A., Tang, J., Xu, L., Huth, A.G.: A natural language fmri dataset for voxelwise encoding models. bioRxiv (2022)
2022
Closest in time.
2022
Closest in time.
Antonello, R.J., Huth, A.: Predictive coding or just feature discovery? an alternative account of why language models fit brain data. Neurobiology of Language, 1–39 (2022)
2022
Closest in time.
Caucheteux, C., King, J.-R.: Brains and algorithms partially converge in natural language processing. Communications biology 5
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
Hazourli, A.: Financialbert - a pretrained language model for financial text mining (2022) https://doi.org/10.13140/RG.2.2.34032.12803
2022
Closest in time.
2022
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.