Fetching the paper…
Reading the bibliography…
We propose the conditional predictive impact (CPI), a consistent and unbiased estimator of the association between one or several features and a given outcome, conditional on a reduced feature set.
Guedj B (2019) A primer on PAC-Bayesian learning, preprint arXiv:1901.05353
1901
Earlier work this paper cites.
Bates S, Candès E, Janson L, Wang W (2019) Metropolized knockoff sampling, preprint arXiv:1903.00434
1903
Earlier work this paper cites.
Fisher RA (1935) The Design of Experiments. Oliver & Boyd, London
1935
Earlier work this paper cites.
Vapnik V, Chervonenkis A (1971) On the uniform convergence of relative frequencies to their probabilities. Theory Probab Appl 16(2):264–280
1971
Earlier work this paper cites.
Sauer N (1972) On the density of families of sets. J Comb Theory Ser A 13(1):145–147
1972
Earlier work this paper cites.
Shelah S (1972) A combinatorial problem: stability and orders for models and theories in infinitariy languages. Pac J Math 41(1):247–261
1972
Earlier work this paper cites.
Harrison D, Rubinfeld DL (1978) Hedonic housing prices and the demand for clean air. J Environ Econ Manag 5(1):81–102
1978
Earlier work this paper cites.
Holm S (1979) A simple sequentially rejective multiple test procedure. Scand J Stat 6(2):65–70
1979
Earlier work this paper cites.
Lindeman RH, Merenda PF, Gold RZ (1980) Introduction to Bivariate and Multivariate Analysis. Longman, Glenview, IL
1980
Earlier work this paper cites.
Pearl J (1988) Probabilistic Reasoning in Intelligent Systems: Networks of Plausible Inference. Morgan Kaufmann, San Mateo, CA
1988
Earlier work this paper cites.
Verma T, Pearl J (1991) Equivalence and synthesis of causal models. In: Proceedings of the Sixth Annual Conference on Uncertainty in Artificial Intelligence, New York, NY, USA, pp 255–270
1991
Earlier work this paper cites.
Benjamini Y, Hochberg Y (1995) Controlling the false discovery rate: a practical and powerful approach to multiple testing. J Royal Stat Soc Ser B Methodol 57(1):289–300
1995
Earlier work this paper cites.
Tibshirani R (1996) Regression shrinkage and selection via the Lasso. J Royal Stat Soc Ser B Methodol 58(1):267–288
1996
Earlier work this paper cites.
Wolpert DH, Macready WG (1997) No free lunch theorems for optimization. IEEE Trans Evol Comput 1(1):67–82
1997
Earlier work this paper cites.
Spirtes P, Glymour CN, Scheines R (2000) Causation, Prediction, and Search, 2nd edn. MIT Press, Cambridge, MA
2000
Earlier work this paper cites.
Benjamini Y, Yekutieli D (2001) The control of the false discovery rate in multiple testing under dependency. Ann Statist 29(4):1165–1188
2001
Earlier work this paper cites.
Breiman L (2001) Random forests. Mach Learn 45(1):1–33
2001
Earlier work this paper cites.
Storey JD (2002) A direct approach to false discovery rates. J Royal Stat Soc Ser B Methodol 64(3):479–498
2002
Earlier work this paper cites.
Venables WN, Ripley BD (2002) Modern Applied Statistics with S, 4th edn. Springer, New York
2002
Earlier work this paper cites.
Guyon I, Elisseeff A (2003) An introduction to variable and feature selection. J Mach Learn Res 3:1157–1182
2003
Earlier work this paper cites.
Sørlie T, Tibshirani R, Parker J, Hastie T, Marron JS, Nobel A, Deng S, Johnsen H, Pesich R, Geisler S, Demeter J, Perou CM, Lønning PE, Brown PO, Børresen-Dale AL, Botstein D (2003) Repeated observation of breast tumor subtypes in independent gene expression data sets. Proc Natl Acad Sci 100(14):8418–8423
2003
Earlier work this paper cites.
Fleuret F (2004) Fast binary feature selection with conditional mutual information. J Mach Learn Res 5:1531–1555
2004
Earlier work this paper cites.
Subramanian A, Tamayo P, Mootha VK, Mukherjee S, Ebert BL, Gillette MA, Paulovich A, Pomeroy SL, Golub TR, Lander ES, Mesirov JP (2005) Gene set enrichment analysis: A knowledge-based approach for interpreting genome-wide expression profiles. Proc Natl Acad Sci 102(43):15545–15550
2005
Earlier work this paper cites.
van der Laan MJ (2006) Statistical inference for variable importance. Int J Biostat 2(1)
2006
Earlier work this paper cites.
Turner NC, Reis-Filho JS (2006) Basal-like breast cancer and the BRCA1 phenotype. Oncogene 25:5846
2006
Earlier work this paper cites.
Grömping U (2007) Estimators of relative importance in linear regression based on variance decomposition. Am Stat 61(2):139–147
2007
Earlier work this paper cites.
Herschkowitz JI, Simin K, Weigman VJ, Mikaelian I, Usary J, Hu Z, Rasmussen KE, Jones LP, Assefnia S, Chandrasekharan S, Backlund MG, Yin Y, Khramtsov AI, Bastein R, Quackenbush J, Glazer RI, Brown PH, Green JE, Kopelovich L, Furth PA, Palazzo JP, Olopade OI, Bernard PS, Churchill GA, Van Dyke T, Perou CM (2007) Identification of conserved gene expression features between murine mammary carcinoma models and human breast tumors. Genome Biol 8(5):R76
2007
Earlier work this paper cites.
Friedman JH, Popescu BE (2008) Predictive learning via rule ensembles. Ann Appl Stat 2(3):916–954
2008
Earlier work this paper cites.
Fukumizu K, Gretton A, Sun X, Schölkopf B (2008) Kernel measures of conditional dependence. In: Advances in Neural Information Processing Systems 20, pp 489–496
2008
Cited alongside, same era.
Strobl C, Boulesteix AL, Kneib T, Augustin T, Zeileis A (2008) Conditional variable importance for random forests. BMC Bioinform 9(1):307
2008
Cited alongside, same era.
Vejmelka M, Paluš M (2008) Inferring the directionality of coupling with conditional mutual information. Phys Rev E 77:026214
2008
Cited alongside, same era.
Koller D, Friedman N (2009) Probabilistic Graphical Models: Principles and Techniques. MIT Press, Cambridge, MA
2009
Cited alongside, same era.
Korb KB, Nicholson AE (2009) Bayesian Artificial Intelligence, 2nd edn. Chapman and Hall/CRC, Boca Raton, FL
2009
Cited alongside, same era.
Ribeiro MT, Singh S, Guestrin C (2016) ”Why Should I Trust You?”: Explaining the Predictions of Any Classifier. In: Proceedings of the 22nd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, ACM, New York, NY, USA, KDD ’16, pp 1135–1144, DOI 10.1145/2939672.2939778
2016
Later among the works it cites.
Dua D, Taniskidou K (2017) UCI Machine Learning Repository. University of California, School of Information and Computer Science, Irvine, CA
2017
Later among the works it cites.
Lei J, G’Sell M, Rinaldo A, Tibshirani RJ, Wasserman L (2018) Distribution-free predictive inference for regression. J Am Stat Assoc 113(523):1094–1111, DOI 10.1080/01621459.2017.1307116
2017
Later among the works it cites.
Lundberg SM, Lee SI (2017) A unified approach to interpreting model predictions. In: Guyon I, Luxburg UV, Bengio S, Wallach H, Fergus R, Vishwanathan S, Garnett R (eds) Advances in Neural Information Processing Systems 30, Curran Associates, Inc., pp 4765–4774, URL http://papers.nips.cc/paper/7062-a-unified-approach-to-interpreting-model-predictions.pdf
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lim E, Vaillant F, Wu D, Forrest NC, Pal B, Hart AH, Asselin-Labat ML, Gyorki DE, Ward T, Partanen A, Feleppa F, Huschtscha LI, Thorne HJ, KConFab, Fox SB, Yan M, French JD, Brown MA, Smyth GK, Visvader JE, Lindeman GJ (2009) Aberrant luminal progenitors as the candidate target population for basal tumor development in BRCA1 mutation carriers. Nat Med 15:907
2009
Cited alongside, same era.
Maathuis MH, Kalisch M, Bühlmann P (2009) Estimating high-dimensional intervention effects from observational data. Ann Statist 37(6A):3133–3164
2009
Cited alongside, same era.
Rouder JN, Speckman PL, Sun D, Morey RD, Iverson G (2009) Bayesian t t -tests for accepting and rejecting the null hypothesis. Psychon Bull Rev 16(2):225–237
2009
Cited alongside, same era.
Wetzels R, Raaijmakers JG, Jakab E, Wagenmakers EJ (2009) How to quantify support for and against the null hypothesis: A flexible winbugs implementation of a default bayesian t t -test. Psychon Bull Rev 16(4):752–760
2009
Cited alongside, same era.
Friedman JH, Hastie T, Tibshirani R (2010) Regularization paths for generalized linear models via coordinate descent. J Stat Softw 33
2010
Cited alongside, same era.
Kursa MB, Rudnicki WR (2010) Feature selection with the boruta package. J Stat Softw 36(11)
2010
Cited alongside, same era.
Meinshausen N, Bühlmann P (2010) Stability selection. J Royal Stat Soc Ser B Methodol 72(4):417–473
2010
Cited alongside, same era.
2017
Later among the works it cites.
Shrikumar A, Greenside P, Kundaje A (2017) Learning important features through propagating activation differences. In: Proceedings of the 34th International Conference on Machine Learning - Volume 70, JMLR.org, ICML’17, pp 3145–3153, URL http://dl.acm.org/citation.cfm?id=3305890.3306006
2017
Later among the works it cites.
Sundararajan M, Taly A, Yan Q (2017) Axiomatic Attribution for Deep Networks. In: Proceedings of the 34th International Conference on Machine Learning - Volume 70, JMLR.org, ICML’17, pp 3319–3328
2017
Later among the works it cites.
Williamson B, Gilbert P, Simon N, Carone M (2017) Nonparametric variable importance assessment using machine learning techniques, preprint UW Biostatistics Working Paper 422
2017
Later among the works it cites.
Wright MN, Ziegler A (2017) ranger: A fast implementation of random forests for high dimensional data in C++ and R. J Stat Softw 77(1)
2017
Later among the works it cites.
Candès E, Fan Y, Janson L, Lv J (2018) Panning for gold: ‘model-X’ knockoffs for high dimensional controlled variable selection. J Royal Stat Soc Ser B Methodol 80(3):551–577
2018
Later among the works it cites.
Feng J, Williamson B, Simon N, Carone M (2018) Nonparametric variable importance using an augmented neural network with multi-task learning. In: Proceedings of the 35th International Conference on Machine Learning, Stockholm, Sweden, Proceedings of Machine Learning Research, vol 80, pp 1496–1505
2018
Later among the works it cites.
Hubbard AE, Kennedy CJ, van der Laan MJ (2018) Data-adaptive target parameters. In: Targeted Learning in Data Science, Springer, New York, chap 9, pp 125–142
2018
Later among the works it cites.
van der Laan MJ, Rose S (eds) (2018) Targeted Learning in Data Science. Springer, New York
2018
Later among the works it cites.
Meyer D, Dimitriadou E, Hornik K, Weingessel A, Leisch F (2018) e1071: Misc Functions of the Department of Statistics, Probability Theory Group, TU Wien. URL https://CRAN.R-project.org/package=e1071 , R package version 1.7-0
2018
Later among the works it cites.
Patterson E, Sesia M (2018) knockoff: The Knockoff Filter for Controlled Variable Selection. URL https://web.stanford.edu/group/candes/knockoffs/index.html , R package version 0.3.2
2018
Later among the works it cites.
R Core Team (2018) R: A Language and Environment for Statistical Computing. Vienna, Austria
2018
Later among the works it cites.
Romano Y, Sesia M, Candès EJ (2018) Deep knockoffs, preprint arXiv:1811.06687
2018
Later among the works it cites.
Sesia M, Sabatti C, Candès EJ (2018) Gene hunting with hidden Markov model knockoffs. Biometrika 106(1):1–18
2018
Later among the works it cites.
Strobl EV, Kun Z, Shyam V (2018) Approximate Kernel-Based Conditional Independence Tests for Fast Non-Parametric Causal Discovery. Journal of Causal Inference 7(1), DOI 10.1515/jci-2018-0017
2018
Later among the works it cites.
Wachter S, Mittelstadt B, Russell C (2018) Counterfactual explanations without opening the black box: automated decisions and the GDPR. Harvard Journal of Law and Technology 31(2):841–887
2018
Later among the works it cites.
Berrett TB, Wang Y, Barber RF, Samworth RJ (2019) The conditional permutation test for independence while controlling for confounders. J Royal Stat Soc Ser B Methodol DOI 10.1111/rssb.12340
2019
Closest in time.
Fisher A, Rudin C, Dominici F (2019) All models are wrong, but many are useful: Learning a variable’s importance by studying an entire class of prediction models simultaneously. J Mach Learn Res 20(177):1–81, URL http://jmlr.org/papers/v20/18-760.html
2019
Closest in time.
Jordon J, Yoon J, van der Schaar M (2019) KnockoffGAN: Generating knockoffs for feature selection using generative adversarial networks. In: International Conference on Learning Representations, New Orleans, USA
2019
Closest in time.
Kuhn M, Johnson K (2019) Feature Engineering and Selection: A Practical Approach for Predictive Models. Chapman and Hall/CRC, Boca Raton, FL
2019
Closest in time.
Rinaldo A, Wasserman L, G’Sell M (2019) Bootstrapping and sample splitting for high-dimensional, assumption-lean inference. Ann Statist 47(6):3438–3469, DOI 10.1214/18-AOS1784
2019
Closest in time.
Shah R, Peters J (2020) The Hardness of Conditional Independence Testing and the Generalised Covariance Measure. Annals of Statistics DOI https://doi.org/10.17863/CAM.45267
2020
Closest in time.
Steinke T, Zakynthinou L (2020) Reasoning about generalization via conditional mutual information. In: Abernethy J, Agarwal S (eds) Proceedings of Thirty Third Conference on Learning Theory, Proceedings of Machine Learning Research, vol 125, pp 3437–3452
2020
Closest in time.
Martínez Sotoca J, Pla F (2010) Supervised feature selection by clustering using conditional mutual information-based distances. Pattern Recognition 43(6):2068–2081
2081
Closest in time.
Barber RF, Candès EJ (2015) Controlling the false discovery rate via knockoffs. Ann Statist 43(5):2055–2085
2085
Closest in time.