Fetching the paper…
Reading the bibliography…
Tree-based machine learning models such as random forests, decision trees, and gradient boosted trees are the most popular non-linear predictive models used in practice today, yet comparatively little attention has been paid to explaining their predictions.
“Comparison of risk prediction using the CKD-EPI equation and the MDRD study equation for estimated glomerular filtration rate”
Kunihiro Matsushita et al · 1951
Earlier work this paper cites.
“A value for n-person games”
Lloyd Shapley · 1953
Earlier work this paper cites.
“Classification and regression trees”
Leo Breiman, Jerome Friedman, Charles Stone and Richard Olshen · 1984
Earlier work this paper cites.
“Monotonic solutions of cooperative games”
H Young · 1985
Earlier work this paper cites.
“The Shapley value: essays in honor of Lloyd S. Shapley”
Alvin Roth · 1988
Earlier work this paper cites.
“Prognostic value of serum creatinine and effect of treatment of hypertension on renal function. Results from the hypertension detection and follow-up program. The Hypertension Detection and Follow-up Program Cooperative Group.”
Neil Shulman et al · 1989
Earlier work this paper cites.
“Renal function change in hypertensive members of the Multiple Risk Factor Intervention Trial: racial and treatment effects”
W Walker et al · 1992
Earlier work this paper cites.
“Body mass index, weight change, and risk of mobility disability in middle-aged and older women: the epidemiologic follow-up study of NHANES I”
Lenore Launer, Tamara Harris, Catherine Rumpel and Jennifer Madans · 1994
Earlier work this paper cites.
“Early predictors of 15-year end-stage renal disease in hypertensive patients”
H Perry et al · 1995
Earlier work this paper cites.
“Plan and operation of the NHANES I Epidemiologic Followup Study, 1992”, 1997
Christine Cox et al · 1997
Earlier work this paper cites.
“Serum uric acid and cardiovascular mortality: the NHANES I epidemiologic follow-up study, 1971-1992”
Jing Fang and Michael Alderman · 2000
Earlier work this paper cites.
“Time-dependent ROC curves for censored survival data and a diagnostic marker”
Patrick Heagerty, Thomas Lumley and Margaret Pepe · 2000
Earlier work this paper cites.
“Random forests”
Leo Breiman · 2001
Earlier work this paper cites.
“Greedy function approximation: a gradient boosting machine”
Jerome Friedman · 2001
Earlier work this paper cites.
“The elements of statistical learning”
Jerome Friedman, Trevor Hastie and Robert Tibshirani · 2001
Earlier work this paper cites.
“Analysis of regression in game theory approach”
Stan Lipovetsky and Michael Conklin · 2001
Earlier work this paper cites.
“NP-completeness for calculating power indices of weighted majority games”
Yasuko Matsui and Tomomi Matsui · 2001
Earlier work this paper cites.
“Effects of blood pressure level on progression of diabetic nephropathy: results from the RENAAL study”
George Bakris et al · 2003
Earlier work this paper cites.
“Repeated observation of breast tumor subtypes in independent gene expression data sets”
Therese Sørlie et al · 2003
Earlier work this paper cites.
“Screening large-scale association study data: exploiting interactions using random forests”
Kathryn Lunetta, L Hayward, Jonathan Segal and Paul Van · 2004
Earlier work this paper cites.
“Feature deduction and ensemble design of intrusion detection systems”
S Chebrolu, A Abraham and J Thomas · 2005
Earlier work this paper cites.
“The effect of a lower target blood pressure on the progression of kidney disease: long-term follow-up of the modification of diet in renal disease study”
Mark Sarnak et al · 2005
Earlier work this paper cites.
“Axiomatic characterizations of probabilistic and cardinal-probabilistic interaction indices”
Katsushige Fujimoto, Ivan Kojadinovic and Jean-Luc Marichal · 2006
Earlier work this paper cites.
“Variable importance in binary regression trees and forests”
Hemant Ishwaran · 2007
Earlier work this paper cites.
“Bias in random forest variable importance measures: Illustrations, sources and a solution”
Carolin Strobl, Anne-Laure Boulesteix, Achim Zeileis and Torsten Hothorn · 2007
Earlier work this paper cites.
“A bias correction algorithm for the Gini variable importance measure in classification trees”
Marco Sandri and Paola Zuccolotto · 2008
Cited alongside, same era.
“Conditional variable importance for random forests”
C Strobl et al · 2008
Cited alongside, same era.
“A random forest approach to the detection of epistatic interactions in case-control studies”
Rui Jiang, Wanwan Tang, Xuebing Wu and Wenhui Fu · 2009
Cited alongside, same era.
“Chronic Renal Insufficiency Cohort (CRIC) Study: baseline characteristics and associations with kidney function”
James Lash et al · 2009
Cited alongside, same era.
“How to explain individual classification decisions”
David Baehrens et al · 2010
Cited alongside, same era.
“Variable selection using random forests”
Robin Genuer, Jean-Michel Poggi and Christine Tuleau-Malot · 2010
“Not Just a Black Box: Learning Important Features Through Propagating Activation Differences”
Avanti Shrikumar, Peyton Greenside, Anna Shcherbina and Anshul Kundaje · 2016
Later among the works it cites.
“Multinational assessment of accuracy of equations for predicting risk of kidney failure: a meta-analysis”
Navdeep Tangri et al · 2016
Later among the works it cites.
“Association between monocyte count and risk of incident CKD and progression to ESRD”
Benjamin Bowe et al · 2017
Later among the works it cites.
“Effects of intensive BP control in CKD”
Alfred Cheung et al · 2017
Later among the works it cites.
“White blood cell count predicts the odds of kidney function decline in a Chinese community-based population”
Fangfang Fan et al · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“Inferring regulatory networks from expression data using tree-based methods”
A Irrthum, L Wehenkel and P Geurts · 2010
Cited alongside, same era.
“The identification of Parkinson’s disease subtypes using cluster analysis: a systematic review”
Stephanie van Rooden et al · 2010
Cited alongside, same era.
“Empirical comparison of tree ensemble variable importance measures”
Lidia Auret and Chris Aldrich · 2011
Cited alongside, same era.
“Scikit-learn: Machine learning in Python”
Fabian Pedregosa et al · 2011
Cited alongside, same era.
“Comparison of lesion formation between contact force-guided and non-guided circumferential pulmonary vein isolation: a prospective, randomized study.”
Masaomi Kimura et al · 2014
Cited alongside, same era.
“Understanding random forests: From theory to practice”
Gilles Louppe · 2014
Cited alongside, same era.
Kaggle · 2017
Later among the works it cites.
“Variable Importance using Decision Trees”
Jalil Kazemitabar, Arash Amini, Adam Bloniarz and Ameet Talwalkar · 2017
Later among the works it cites.
“Lightgbm: A highly efficient gradient boosting decision tree”
Guolin Ke et al · 2017
Later among the works it cites.
“Learning how to explain neural networks: PatternNet and PatternAttribution”
Pieter-Jan Kindermans et al · 2017
Later among the works it cites.
“A Unified Approach to Interpreting Model Predictions”
Scott Lundberg and Su-In Lee · 2017
Later among the works it cites.
“Axiomatic attribution for deep networks”
Mukund Sundararajan, Ankur Taly and Qiqi Yan · 2017
Later among the works it cites.
“Rules of Machine Learning: Best Practices for ML Engineering”
Martin Zinkevich · 2017
Later among the works it cites.
“Towards better understanding of gradient-based attribution methods for Deep Neural Networks”
Marco Ancona, Enea Ceolini, Cengiz Oztireli and Markus Gross · 2018
Later among the works it cites.
“Evaluating feature importance estimates”
Sara Hooker, Dumitru Erhan, Pieter-Jan Kindermans and Been Kim · 2018
Later among the works it cites.
“DeepSurv: personalized treatment recommender system using a Cox proportional hazards deep neural network”
Jared Katzman et al · 2018
Later among the works it cites.
“Explainable machine learning predictions to help anesthesiologists prevent hypoxemia during surgery”
Scott Lundberg et al · 2018
Later among the works it cites.
“Hematuria as a risk factor for progression of chronic kidney disease and death: findings from the Chronic Renal Insufficiency Cohort (CRIC) Study”
Paula Orlandi et al · 2018
Later among the works it cites.
“Model Agnostic Supervised Local Explanations”
Gregory Plumb, Denali Molitor and Ameet Talwalkar · 2018
Later among the works it cites.
“CatBoost: unbiased boosting with categorical features”
Liudmila Prokhorenkova et al · 2018
Later among the works it cites.
“Anchors: High-precision model-agnostic explanations”
Marco Ribeiro, Sameer Singh and Carlos Guestrin · 2018
Later among the works it cites.
“Clinical Decision Support in the Era of Artificial Intelligence”
Edward Shortliffe and Martin Sepúlveda · 2018
Later among the works it cites.
“Feature selection for ranking using boosted trees”
Feng Pan et al · 2028
Closest in time.
“The association of blood pressure levels and change in renal function in hypertensive and nonhypertensive subjects”
Steven Rosansky, Donald Hoover, Lisa King and James Gibson · 2076
Closest in time.
“Association of estimated glomerular filtration rate and albuminuria with all-cause and cardiovascular mortality in general population cohorts: a collaborative meta-analysis”
Chronic Consortium · 2081
Closest in time.
“Urinary creatinine excretion, bioelectrical impedance analysis, and clinical outcomes in patients with CKD: the CRIC study”
F Wilson et al · 2095
Closest in time.