Fetching the paper…
Reading the bibliography…
Prediction-powered inference is a framework for performing valid statistical inference when an experimental dataset is supplemented with predictions from a machine-learning system.
“Asymptotic minimax character of the sample distribution function and of the classical multinomial estimator”
Aryeh Dvoretzky, Jack Kiefer and Jacob Wolfowitz · 1956
Earlier work this paper cites.
“On the central limit theorem for samples from a finite population”
Paul Erdős · 1959
Earlier work this paper cites.
“Some results on generalized difference estimation and generalized regression estimation for finite populations”
Claes Cassel, Carl Särndal and Jan Wretman · 1976
Earlier work this paper cites.
“Regression quantiles”
Roger Koenker and Gilbert Bassett · 1978
Earlier work this paper cites.
“Sampling from a finite population. A remainder term estimate”
Thomas Höglund · 1978
Earlier work this paper cites.
“Using least squares to approximate unknown regression functions”
Halbert White · 1980
Earlier work this paper cites.
“A heteroskedasticity-consistent covariance matrix estimator and a direct test for heteroskedasticity”
Halbert White · 1980
Earlier work this paper cites.
“The tight constant in the Dvoretzky-Kiefer-Wolfowitz inequality”
Pascal Massart · 1990
Earlier work this paper cites.
“Model Assisted Survey Sampling”
Carl-Erik Särndal, Bengt Swensson and Jan Wretman · 1992
Earlier work this paper cites.
“Inference using surrogate outcome data and a validation sample”
Margaret Pepe · 1992
Earlier work this paper cites.
“Estimation of regression coefficients when some regressors are not always observed”
James Robins, Andrea Rotnitzky and Lue Zhao · 1994
Earlier work this paper cites.
“Semiparametric efficiency in multivariate regression models with missing data”
James Robins and Andrea Rotnitzky · 1995
Earlier work this paper cites.
“The sloan digital sky survey: Technical summary”
Donald York, J Adelman, John Anderson, Scott Anderson, James Annis, Neta Bahcall, JA Bakken, Robert Barkhouser, Steven Bastian and Eileen Berman · 2000
Earlier work this paper cites.
“A Model-Calibration Approach to Using Complete Auxiliary Information From Survey Data”
Changbao Wu and Randy Sitter · 2001
Earlier work this paper cites.
“Information recovery in a study with surrogate endpoints”
Song Chen, Denis Leung and Jing Qin · 2003
Earlier work this paper cites.
“An automated submersible flow cytometer for analyzing pico-and nanophytoplankton: FlowCytobot”
Robert Olson, Alexi Shalapyonok and Heidi Sosik · 2003
Earlier work this paper cites.
“PDB file parser and structure class implemented in Python”
Thomas Hamelryck and Bernard Manderick · 2003
Earlier work this paper cites.
“Semiparametric efficient estimation for the auxiliary outcome problem with the conditional mean model”
Jinbo Chen and Norman Breslow · 2004
Earlier work this paper cites.
“The importance of intrinsic disorder for protein phosphorylation”
Lilia Iakoucheva, Predrag Radivojac, Celeste Brown, Timothy O’Connor, Jason Sikes, Zoran Obradovic and A Dunker · 2004
Earlier work this paper cites.
“Measurement error models with auxiliary data”
Xiaohong Chen, Han Hong and Elie Tamer · 2005
Earlier work this paper cites.
“Semi-supervised learning literature survey”
Xiaojin Zhu · 2005
Earlier work this paper cites.
“A revisit of semiparametric regression models with missing data”
Menggang Yu and Bin Nan · 2006
Cited alongside, same era.
“Measurement Error in Nonlinear Models: A Modern Perspective”
Raymond Carroll, David Ruppert, Leonard Stefanski and Ciprian Crainiceanu · 2006
Cited alongside, same era.
“Statistical Analysis of Semi-Supervised Regression”
Larry Wasserman and John Lafferty · 2007
Cited alongside, same era.
“Introduction to semi-supervised learning”
Xiaojin Zhu and Andrew Goldberg · 2009
Cited alongside, same era.
“Galaxy Zoo 2: detailed morphological classifications for 304 122 galaxies from the Sloan Digital Sky Survey”
Kyle Willett, Chris Lintott, Steven Bamford, Karen Masters, Brooke Simmons, Kevin Casteels, Edward Edmondson, Lucy Fortson, Sugata Kaviraj and William Keel · 2013
Cited alongside, same era.
“Global, 30-m resolution continuous fields of tree cover: Landsat-based rescaling of MODIS vegetation continuous fields with lidar-based estimates of error”
“Satellite-based estimates reveal widespread forest degradation in the Amazon”
Eric Bullock, Curtis Woodcock, Carlos Souza and Pontus Olofsson · 2020
Later among the works it cites.
“Estimating means of bounded random variables by betting”
Ian Waudby-Smith and Aaditya Ramdas · 2020
Later among the works it cites.
“A short note on learning discrete distributions”
Clément Canonne · 2020
Later among the works it cites.
“Highly accurate protein structure prediction with AlphaFold”
John Jumper, Richard Evans, Alexander Pritzel, Tim Green, Michael Figurnov, Olaf Ronneberger, Kathryn Tunyasuvunakool, Russ Bates, Augustin Žídek and Anna Potapenko · 2021
Later among the works it cites.
“Highly accurate protein structure prediction for the human proteome”
Kathryn Tunyasuvunakool, Jonas Adler, Zachary Wu, Tim Green, Michal Zielinski, Augustin Žídek, Alex Bridgland, Andrew Cowie, Clemens Meyer and Agata Laydon · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Joseph Sexton, Xiao-Peng Song, Min Feng, Praveen Noojipady, Anupam Anand, Chengquan Huang, Do-Hyung Kim, Kathrine Collins, Saurabh Channan and Charlene DiMiceli · 2013
Cited alongside, same era.
“Adam: A method for stochastic optimization”
Diederik Kingma and Jimmy Ba · 2014
Cited alongside, same era.
“UniProt: a hub for protein information”
UniProt Consortium · 2015
Cited alongside, same era.
Eric Orenstein, Oscar Beijbom, Emily Peacock and Heidi Sosik · 2015
Cited alongside, same era.
“Xgboost: A scalable tree boosting system”
Tianqi Chen and Carlos Guestrin · 2016
Cited alongside, same era.
“Deep residual learning for image recognition”
Kaiming He, Xiangyu Zhang, Shaoqing Ren and Jian Sun · 2016
Cited alongside, same era.
“Model-assisted survey estimation with modern prediction techniques”
F Breidt and Jean Opsomer · 2017
Cited alongside, same era.
Later among the works it cites.
“Semi-Supervised Linear Regression”
David Azriel, Lawrence. Brown, Michael Sklar, Richard Berk, Andreas Buja and Linda Zhao · 2021
Later among the works it cites.
“Surrogate assisted semi-supervised inference for high dimensional risk prediction”
Jue Hou, Zijian Guo and Tianxi Cai · 2021
Later among the works it cites.
“Predicting with proxies: Transfer learning in high dimension”
Hamsa Bastani · 2021
Later among the works it cites.
“Learning across bandits in high dimension via robust statistics”
Kan Xu and Hamsa Bastani · 2021
Later among the works it cites.
“Retiring adult: New datasets for fair machine learning”
Frances Ding, Moritz Hardt, John Miller and Ludwig Schmidt · 2021
Later among the works it cites.
“The structural context of posttranslational modifications at a proteome-wide scale”
Isabell Bludau, Sander Willems, Wen-Feng Zeng, Maximilian Strauss, Fynn Hansen, Maria Tanzer, Ozge Karayel, Brenda Schulman and Matthias Mann · 2022
Later among the works it cites.
“Semi-Supervised Quantile Estimation: Robust and Efficient Inference in High Dimensional Settings”
Abhishek Chakrabortty, Guorong Dai and Raymond Carroll · 2022
Later among the works it cites.
Abhishek Chakrabortty, Guorong Dai and Eric Tchetgen · 2022
Later among the works it cites.
“High-dimensional semi-supervised learning: in search of optimal inference of the mean”
Yuqian Zhang and Jelena Bradic · 2022
Later among the works it cites.
“Transfer learning under high-dimensional generalized linear models”
Ye Tian and Yang Feng · 2022
Later among the works it cites.
“On transfer learning in functional linear regression”
Haotian Lin and Matthew Reimherr · 2022
Later among the works it cites.
“Exponential Families in Theory and Practice”
Bradley Efron · 2022
Later among the works it cites.
“The evolution, evolvability and engineering of gene regulatory DNA”
Eeshit Vaishnav, Carl de Boer, Jennifer Molinet, Moran Yassour, Lin Fan, Xian Adiconis, Dawn Thompson, Joshua Levin, Francisco Cubillos and Aviv Regev · 2022
Later among the works it cites.
“Clustering predicted structures at the scale of the known protein universe”
Inigo Barrio-Hernandez, Jingi Yeo, Jürgen Jänes, Tanita Wein, Mihaly Varadi, Sameer Velankar, Pedro Beltrao and Martin Steinegger · 2023
Closest in time.
“A General M-estimation Theory in Semi-Supervised Framework”
Shanshan Song, Yuanyuan Lin and Yong Zhou · 2023
Closest in time.
“ ppi-py
Anastasios Angelopoulos, Stephen Bates, Clara Fannjiang, Michael Jordan and Tijana Zrnic · 2023
Closest in time.