Fetching the paper…
Reading the bibliography…
In computational social science (CSS), researchers analyze documents to explain social and political phenomena.
Sentence-bert: Sentence embeddings using siamese bert-networks
Nils Reimers and Iryna Gurevych · 1908
Earlier work this paper cites.
A three-population model for sequential screening for bacteriuria
Paul S Levy and Edward H Kass · 1970
Earlier work this paper cites.
Estimation of Regression Coefficients When Some Regressors Are Not Always Observed
James M Robins, Andrea Rotnitzky, and Lue Ping Zhao · 1994
Earlier work this paper cites.
Semiparametric efficiency in multivariate regression models with missing data
James M Robins and Andrea Rotnitzky · 1995
Earlier work this paper cites.
Misclassification of the dependent variable in a discrete-response setting
Jerry A Hausman, Jason Abrevaya, and Fiona M Scott-Morton · 1998
Earlier work this paper cites.
Probabilistic outputs for support vector machines and comparisons to regularized likelihood methods
John Platt · 1999
Earlier work this paper cites.
Asymptotic Statistics , volume 3
Aad W van der Vaart · 2000
Earlier work this paper cites.
Unified methods for censored longitudinal data and causality
Mark J Laan and James M Robins · 2003
Earlier work this paper cites.
Gender bias in coreference resolution: Evaluation and debiasing methods
Jieyu Zhao, Tianlu Wang, Mark Yatskar, Vicente Ordonez, and Kai-Wei Chang · 2003
Earlier work this paper cites.
Counting positives accurately despite inaccurate classification
George Forman · 2005
Earlier work this paper cites.
Congressional bills project
E Scott Adler and John Wilkerson · 2006
Earlier work this paper cites.
Semiparametric theory and missing data
Anastasios A Tsiatis · 2006
Earlier work this paper cites.
A method of automated nonparametric content analysis for social science
Daniel J Hopkins and Gary King · 2010
Earlier work this paper cites.
Aggregative quantification for regression
Antonio Bella, Cesar Ferri, José Hernández-Orallo, and María José Ramírez-Quintana · 2014
Earlier work this paper cites.
Double-robust methods
Andrea Rotnitzky and Stijn Vansteelandt · 2014
Earlier work this paper cites.
A review on quantification learning
Pablo González, Alberto Castaño, Nitesh V Chawla, and Juan José Del Coz · 2017
Earlier work this paper cites.
The importance of calibration for estimating proportions from annotations
Dallas Card and Noah A Smith · 2018
Earlier work this paper cites.
Efficient and adaptive linear regression in semi-supervised settings
Abhishek Chakrabortty and Tianxi Cai · 2018
Cited alongside, same era.
Uncertainty-aware generative models for inferring document class prevalence
Katherine Keith and Brendan O’Connor · 2018
Cited alongside, same era.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Cited alongside, same era.
On the role of surrogates in the efficient estimation of treatment effects with limited outcome data
Nathan Kallus and Xiaojie Mao · 2020
Cited alongside, same era.
Mpnet: Masked and permuted pre-training for language understanding
Kaitao Song, Xu Tan, Tao Qin, Jianfeng Lu, and Tie-Yan Liu · 2020
Cited alongside, same era.
Text as data: A new framework for machine learning and the social sciences
Justin Grimmer, Margaret E Roberts, and Brandon M Stewart · 2022
Later among the works it cites.
Semiparametric doubly robust targeted double machine learning: a review
Edward H Kennedy · 2022
Later among the works it cites.
Testing causal theories with learned proxies
Dean Knox, Christopher Lucas, and Wendy K Tam Cho · 2022
Later among the works it cites.
How to train your stochastic parrot: Large language models for political texts
Joseph T Ornstein, Elise N Blasingame, and Jake S Truscott · 2022
Later among the works it cites.
grf: Generalized Random Forests , 2022
Julie Tibshirani, Susan Athey, Erik Sverdrup, and Stefan Wager · 2022
Later among the works it cites.
Assumption-lean Inference for Generalised Linear Model Parameters
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Siruo Wang, Tyler H McCormick, and Jeffrey T Leek · 2020
Cited alongside, same era.
On the dangers of stochastic parrots: Can language models be too big?
Emily M Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell · 2021
Cited alongside, same era.
Machine learning predictions as regression covariates
Christian Fong and Matthew Tyler · 2021
Cited alongside, same era.
True few-shot learning with language models
Ethan Perez, Douwe Kiela, and Kyunghyun Cho · 2021
Cited alongside, same era.
How using machine learning classification as a variable in regression leads to attenuation bias and what to do about it
Han Zhang · 2021
Cited alongside, same era.
Calibrate before use: Improving few-shot performance of language models
Zihao Zhao, Eric Wallace, Shi Feng, Dan Klein, and Sameer Singh · 2021
Cited alongside, same era.
A general framework for treatment effect estimation in semi-supervised and high dimensional settings
Abhishek Chakrabortty, Guorong Dai, and Eric Tchetgen Tchetgen · 2022
Cited alongside, same era.
Stijn Vansteelandt and Oliver Dukes · 2022
Later among the works it cites.
Prediction-powered inference
Anastasios N. Angelopoulos, Stephen Bates, Clara Fannjiang, Michael I. Jordan, and Tijana Zrnic · 2023
Closest in time.
ChatGPT outperforms crowd workers for text-annotation tasks
Fabrizio Gilardi, Meysam Alizadeh, and Maël Kubli · 2023
Closest in time.
An improved method of automated nonparametric content analysis for social science
Connor T Jerzak, Gary King, and Anton Strezhnev · 2023
Closest in time.
Statistical analysis with machine learning predicted variables
Hiroto Katsumata and Soichiro Yamauchi · 2023
Closest in time.
Voteview: Congressional roll-call votes database
Jeffrey B Lewis, Keith Poole, Howard Rosenthal, Adam Boche, Aaron Rudkin, and Luke Sonnet · 2023
Closest in time.
Reagan Mozer and Luke Miratrix · 2023
Closest in time.
Petter Törnberg · 2023
Closest in time.
Confronting core issues: A critical test of attitude polarization
Yamil Velez and Patrick Liu · 2023
Closest in time.
Large language models can be used to estimate the latent positions of politicians
Patrick Y. Wu, Jonathan Nagler, Joshua A. Tucker, and Solomon Messing · 2023
Closest in time.
Can large language models transform computational social science?
Caleb Ziems, William Held, Omar Shaikh, Jiaao Chen, Zhehao Zhang, and Diyi Yang · 2023
Closest in time.