Fetching the paper…
Reading the bibliography…
Missing values pose a persistent challenge in modern data science.
[author] Rubin, Donald B.D. B. (1976). Inference and missing data. Biometrika 63 581-592
1976
Earlier work this paper cites.
[author] Little, Roderick J. A.R. J. A. and Donald, Rubin B.R. B. (1986). Statistical Analysis with Missing Data. John Wiley & Sons, Inc
1986
Earlier work this paper cites.
[author] Little, Roderick J. A.R. J. A. (1993). Pattern-Mixture Models for Multivariate Incomplete Data. Journal of the American Statistical Association 88 125-134
1993
Earlier work this paper cites.
[author] Ren, BoyuB., Lipsitz, Stuart RS. R., Weiss, Roger DR. D. and Fitzmaurice, Garrett MG. M. (2023). Multiple imputation for non-monotone missing not at random data using the no self-censoring model. Stat Methods Med Res 32 1973–1993
1993
Earlier work this paper cites.
[author] Schafer, Joseph LJ. L. (1997). Analysis of incomplete multivariate data. Chapman and Hall/CRC
1997
Earlier work this paper cites.
[author] Ibrahim, Joseph G.J. G., Lipsitz, Stuart R.S. R. and Chen, Ming-HuiM.-H. (1999). Missing covariates in generalized linear models when the missing data mechanism is non-ignorable. Journal of the Royal Statistical Society: Series B (Statistical Methodology) 61 173-190
1999
Earlier work this paper cites.
[author] Breiman, LeoL. (2001). Random forests. Machine learning 45 5–32
2001
Earlier work this paper cites.
[author] Székely, Gabor J.G. J. (2003). E-statistics: the energy of statistical samples Technical Report No. 05, Bowling Green State University, Department of Mathematics and Statistics
2003
Earlier work this paper cites.
[author] Nelsen, Roger B.R. B. (2006). An Introduction to Copulas, 2nd ed. Springer Series in Statistics. Springer, New York
2006
Earlier work this paper cites.
[author] Potthoff, Richard FR. F., Tudor, Gail EG. E., Pieper, Karen SK. S. and Hasselblad, VicV. (2006). Can one assess whether missing data are missing at random in medical studies? Stat Methods Med Res 15 213–234
2006
Earlier work this paper cites.
[author] Gneiting, TilmannT. and Raftery, Adrian EA. E. (2007). Strictly Proper Scoring Rules, Prediction, and Estimation. Journal of the American Statistical Association 102 359-378
2007
Earlier work this paper cites.
[author] Van Buuren, StefS. (2007). Multiple imputation of discrete and continuous data by fully conditional specification. Stat Methods Med Res 16 219–242
2007
Earlier work this paper cites.
[author] Molenberghs, GeertG., Beunckens, CarolineC., Sotto, CristinaC. and Kenward, Michael G.M. G. (2008). Every Missingness Not at Random Model Has a Missingness at Random Counterpart with Equal Fit. Journal of the Royal Statistical Society. Series B (Statistical Methodology) 70 371–388
2008
Earlier work this paper cites.
[author] Burgette, Lane F.L. F. and Reiter, Jerome P.J. P. (2010). Multiple Imputation for Missing Data via Sequential Regression Trees. American Journal of Epidemiology 172 1070-1076
2010
Earlier work this paper cites.
[author] Van Buuren, StefS. and Groothuis-Oudshoorn, KarinK. (2011). mice: Multivariate Imputation by Chained Equations in R. Journal of Statistical Software 45 1-67
2011
Earlier work this paper cites.
[author] Li, LinglingL., Shen, ChangyuC., Li, XiaochunX. and Robins, James MJ. M. (2011). On weighting approaches for missing data. Stat Methods Med Res 22 14–30
2011
Earlier work this paper cites.
[author] Stekhoven, Daniel J.D. J. and Bühlmann, PeterP. (2011). MissForest—non-parametric missing value imputation for mixed-type data. Bioinformatics 28 112-118
2011
Earlier work this paper cites.
[author] Mohan, KarthikaK., Pearl, JudeaJ. and Tian, JinJ. (2013). Graphical Models for Inference with Missing Data. In Advances in Neural Information Processing Systems 26 1277–1285
2013
Earlier work this paper cites.
[author] Seaman, ShaunS., Galati, JohnJ., Jackson, DanD. and Carlin, JohnJ. (2013). What is meant by “Missing at Random”? Statistical Science 28 257-268
2013
Earlier work this paper cites.
[author] Takai, KeijiK. and Kano, YutakaY. (2013). Asymptotic Inference with Incomplete Data. Communications in Statistics - Theory and Methods 42 3174–3190
2013
Earlier work this paper cites.
[author] Waljee, Akbar K.A. K., Mukherjee, AshinA., Singal, Amit G.A. G., Zhang, YiweiY., Warren, JeffreyJ., Balis, UlyssesU., Marrero, JorgeJ., Zhu, JiJ. and Higgins, Peter DrP. D. (2013). Comparison of imputation methods for missing laboratory data in medicine. BMJ open 3
2013
Earlier work this paper cites.
[author] Doove, Lisa L.L. L., Van Buuren, StefS. and Dusseldorp, EliseE. (2014). Recursive partitioning for missing data imputation in the presence of interaction effects. Computational Statistics & Data Analysis 72 92-104
2014
Earlier work this paper cites.
[author] Liu, JingchenJ., Gelman, AndrewA., Hill, JenniferJ., Su, Yu-SungY.-S. and Kropko, JonathanJ. (2014). On the stationary distribution of iterative imputations. Biometrika 1 155–173
2014
Earlier work this paper cites.
[author] Mohan, KarthikaK. and Pearl, JudeaJ. (2014). On the testability of models with missing data. Artificial Intelligence and Statistics 643–650. PMLR
2014
Earlier work this paper cites.
[author] Bonneel, NicolasN., Rabin, JulienJ., Peyré, GabrielG. and Pfister, HanspeterH. (2015). Sliced and Radon Wasserstein Barycenters of Measures. Journal of Mathematical Imaging and Vision 51 22-45
2015
Cited alongside, same era.
[author] Mealli, FabriziaF. and Rubin, Donald B.D. B. (2015). Clarifying missing at random and related definitions, and implications when coupled with exchangeability. Biometrika 102 995-1000
2015
Cited alongside, same era.
[author] Zhu, JianJ. and Raghunathan, Trivellore E.T. E. (2015). Convergence Properties of a Sequential Regression Multiple Imputation Algorithm. Journal of the American Statistical Association 110 1112-1124
2015
Cited alongside, same era.
[author] Lee, Min CherngM. C. and Mitra, RobinR. (2016). Multiply imputing missing values in data sets with mixed measurement scales using a sequence of generalised linear models. Computational Statistics & Data Analysis 95 24-38
2016
Cited alongside, same era.
Bhattacharya, R
2020
Later among the works it cites.
[author] Cantoni, EvaE. and de Luna, XavierX. (2020). Semiparametric inference with missing data: Robustness to outliers and model misspecification. Econometrics and Statistics 16 108-120
2020
Later among the works it cites.
[author] Frahm, GabrielG., Nordhausen, KlausK. and Oja, HannuH. (2020). M-estimation with incomplete and dependent multivariate data. Journal of Multivariate Analysis 176 104569
2020
Later among the works it cites.
[author] Hong, ShangzhiS. and Lynn, Henry S.H. S. (2020). Accuracy of random-forest-based imputation of missing data in the presence of non-normality, non-linearity, and interaction. BMC Medical Research Methodology 20 199
2020
Later among the works it cites.
Muzellec, B
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Shpitser, I
2016
Cited alongside, same era.
[author] Xu, DandanD., Daniels, Michael J.M. J. and Winterstein, Almut G.A. G. (2016). Sequential BART for imputation of missing covariates. Biostatistics 17 589–602
2016
Cited alongside, same era.
Arjovsky, M
2017
Cited alongside, same era.
[author] Tang, FeiF. and Ishwaran, HemantH. (2017). Random Forest Missing Data Algorithms. Stat Anal Data Min 10 363–377
2017
Cited alongside, same era.
Ambrogioni, L
2018
Cited alongside, same era.
[author] Bertsimas, DimitrisD., Pawlowski, ColinC. and Zhuo, Ying DaisyY. D. (2018). From Predictive Methods to Missing Data Imputation: An Optimization Approach. Journal of Machine Learning Research 18 1–39
2018
Cited alongside, same era.
[author] Van Buuren, StefS. (2018). Flexible Imputation of Missing Data. Second Edition. Chapman & Hall/CRC Press
2018
Cited alongside, same era.
[author] Doretti, MarcoM., Geneletti, SaraS. and Stanghellini, ElenaE. (2018). Missing Data: A Unified Taxonomy Guided by Conditional Independence. International Statistical Review 86 189-204
2018
Cited alongside, same era.
[author] Nazábal, AlfredoA., Olmos, Pablo M.P. M., Ghahramani, ZoubinZ. and Valera, IsabelI. (2020). Handling incomplete heterogeneous data using VAEs. Pattern Recognition 107 107501
2020
Later among the works it cites.
Oberst, M
2020
Later among the works it cites.
[author] Qiu, Yeping LinaY. L., Zheng, HongH. and Gevaert, OlivierO. (2020). Genomic data imputation with variational auto-encoders. GigaScience 9 giaa082
2020
Later among the works it cites.
[author] Dong, WeinanW., Fong, Daniel Yee TakD. Y. T., Yoon, Jin-sunJ.-s., Wan, Eric Yuk FaiE. Y. F., Bedford, Laura ElizabethL. E., Tang, Eric Ho ManE. H. M. and Lam, Cindy Lo KuenC. L. K. (2021). Generative adversarial networks for imputing missing data for big data clinical research. BMC Medical Research Methodology 21 78
2021
Later among the works it cites.
[author] Jäger, SebastianS., Allhorn, ArndtA. and Bießmann, FelixF. (2021). A Benchmark for Data Imputation Methods. Frontiers in Big Data 4
2021
Later among the works it cites.
[author] Mohan, KarthikaK. and Pearl, JudeaJ. (2021). Graphical models of processing missing data. Journal of the American Statistical Association 1–16
2021
Later among the works it cites.
[author] Ćevid, DomagojD., Michel, LorisL., Näf, JeffreyJ., Meinshausen, NicolaiN. and Bühlmann, PeterP. (2022). Distributional Random Forests: Heterogeneity Adjustment and Multivariate Distributional Regression. Journal of Machine Learning Research 23 1–79
2022
Later among the works it cites.
[author] Daniel Malinsky, Ilya ShpitserI. S. and Tchetgen, Eric J. TchetgenE. J. T. (2022). Semiparametric Inference for Nonmonotone Missing-Not-at-Random Data: The No Self-Censoring Model. Journal of the American Statistical Association 117 1415–1423
2022
Later among the works it cites.
[author] Deng, GraceG., Han, CuizeC. and Matteson, David S.D. S. (2022). Extended missing data imputation via GANs for ranking applications. Data Mining and Knowledge Discovery 36 1498-1520
2022
Later among the works it cites.
[author] Imaizumi, MasaakiM., Ota, HirofumiH. and Hamaguchi, TakuoT. (2022). Hypothesis Test and Confidence Analysis With Wasserstein Distance on General Dimension. Neural Computation 34 1448-1487
2022
Later among the works it cites.
Rizzo, M
2022
Later among the works it cites.
[author] Wang, ZhenhuaZ., Akande, OlanrewajuO., Poulos, JasonJ. and Li, FanF. (2022). Are deep learning models superior for missing data imputation in surveys? Evidence from an empirical comparison. Survey Methodology 48
2022
Later among the works it cites.
[author] Fang, FangF. and Bao, ShenliaoS. (2023). FragmGAN: Generative adversarial nets for fragmentary data imputation and prediction. Statistical Theory and Related Fields 0 1-14
2023
Later among the works it cites.
[author] Näf, JeffreyJ., Spohn, Meta-LinaM.-L., Michel, LorisL. and Meinshausen, NicolaiN. (2023). Imputation scores. The Annals of Applied Statistics 17 2452 – 2472
2023
Later among the works it cites.
[author] Rabe-Hesketh, SophiaS. and Skrondal, AndersA. (2023). Ignoring Non-ignorable Missingness. Psychometrika 88 31-50
2023
Later among the works it cites.
[author] Shadbahr, T.T., Roberts, M.M., Stanczuk, J.J. et al. (2023). The impact of imputation quality on machine learning classifiers for datasets with missing values. Communications Medicine 3 139
2023
Later among the works it cites.
2024
Closest in time.
[author] Nabi, RaziehR., Bhattacharya, RohitR., Shpitser, IlyaI. and Robins, JamesJ. (2025). Causal and counterfactual views of missing data models. Statistica Sinica
2025
Closest in time.