Fetching the paper…
Reading the bibliography…
Peer prediction mechanisms motivate high-quality feedback with provable guarantees.
A mathematical theory of communication
Claude Elwood Shannon · 1948
Earlier work this paper cites.
Optimal selling strategies under uncertainty for a discriminating monopolist when demands are interdependent
Jacques Crémer and Richard P. McLean · 1985
Earlier work this paper cites.
Experts in uncertainty: opinion and subjective probability in science
Roger Cooke · 1991
Earlier work this paper cites.
Axiomatic characterization of the quadratic scoring rule
Reinhard Selten · 1998
Earlier work this paper cites.
Reputation systems
Paul Resnick, Ko Kuwabara, Richard Zeckhauser, and Eric Friedman · 2000
Earlier work this paper cites.
A bayesian truth serum for subjective data
Drazen Prelec · 2004
Earlier work this paper cites.
Eliciting informative feedback: The peer-prediction method
Nolan Miller, Paul Resnick, and Richard Zeckhauser · 2005
Earlier work this paper cites.
Strictly proper scoring rules, prediction, and estimation
Tilmann Gneiting and Adrian E Raftery · 2007
Earlier work this paper cites.
Crowdsourced judgement elicitation with endogenous proficiency
Anirban Dasgupta and Arpita Ghosh · 2013
Earlier work this paper cites.
A robust bayesian truth serum for non-binary signals
Goran Radanovic and Boi Faltings · 2013
Earlier work this paper cites.
Incentives for truthful information elicitation of continuous signals
Goran Radanovic and Boi Faltings · 2014
Earlier work this paper cites.
Elicitability and knowledge-free elicitation with peer prediction
Peter Zhang and Yiling Chen · 2014
Earlier work this paper cites.
Incentivizing evaluation via limited access to ground truth: Peer-prediction makes things worse
Alice Gao, James R Wright, and Kevin Leyton-Brown · 2016
Earlier work this paper cites.
Informed truthfulness in multi-task peer prediction
Victor Shnayder, Arpit Agarwal, Rafael Frongillo, and David C Parkes · 2016
Earlier work this paper cites.
Reputation and feedback systems in online platform markets
Steven Tadelis · 2016
Earlier work this paper cites.
Peer prediction with heterogeneous users
Arpit Agarwal, Debmalya Mandal, David C Parkes, and Nisarg Shah · 2017
Earlier work this paper cites.
A solution to the single-question crowd wisdom problem
Dražen Prelec, H Sebastian Seung, and John McCoy · 2017
Earlier work this paper cites.
Robust forecast aggregation
Itai Arieli, Yakov Babichenko, and Rann Smorodinsky · 2018
Earlier work this paper cites.
Recognition in terra incognita
Sara Beery, Grant Van Horn, and Pietro Perona · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Earlier work this paper cites.
Reputation inflation
Apostolos Filippas, John Joseph Horton, and Joseph Golden · 2018
Earlier work this paper cites.
Eliciting expertise without verification
Yuqing Kong and Grant Schoenebeck · 2018
Cited alongside, same era.
An information theoretic framework for designing information elicitation mechanisms that reward truth-telling
Yuqing Kong and Grant Schoenebeck · 2019
Cited alongside, same era.
Probing neural network comprehension of natural language arguments
Timothy Niven and Hung-Yu Kao · 2019
Cited alongside, same era.
Extracting the wisdom of crowds when information is shared
Asa B Palley and Jack B Soll · 2019
Cited alongside, same era.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Cited alongside, same era.
Shortcut learning in deep neural networks
The dangers of using large language models for peer review
Tjibbe Donker · 2023
Later among the works it cites.
Fighting reviewer fatigue or amplifying bias? considerations and recommendations for use of chatgpt and other large language models in scholarly peer review
Mohammad Hosseini and Serge Horbach · 2023
Later among the works it cites.
Can large language models provide useful feedback on research papers? a large-scale empirical analysis, 2023
Weixin Liang, Yuhui Zhang, Hancheng Cao, Binglu Wang, Daisy Ding, Xinyu Yang, Kailas Vodrahalli, Siyu He, Daniel Smith, Yian Yin, Daniel McFarland, and James Zou · 2023
Later among the works it cites.
Reviewergpt? an exploratory study on using large language models for paper reviewing, 2023
Ryan Liu and Nihar B. Shah · 2023
Later among the works it cites.
Surrogate scoring rules
Yang Liu, Juntao Wang, and Yiling Chen · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Robert Geirhos, Jörn-Henrik Jacobsen, Claudio Michaelis, Richard Zemel, Wieland Brendel, Matthias Bethge, and Felix A Wichmann · 2020
Cited alongside, same era.
Dominantly truthful multi-task peer prediction with a constant number of tasks
Yuqing Kong · 2020
Cited alongside, same era.
Learning and strongly truthful multi-task peer prediction: A variational approach
Grant Schoenebeck and Fang-Yi Yu · 2020
Cited alongside, same era.
Measurement integrity in peer prediction: A peer assessment case study
Noah Burrell and Grant Schoenebeck · 2021
Cited alongside, same era.
The wisdom of the crowd and higher-order beliefs
Yi-Chun Chen, Manuel Mueller-Frank, and Mallesh M Pai · 2021
Cited alongside, same era.
Information elicitation from rowdy crowds
Grant Schoenebeck, Fang-Yi Yu, and Yichi Zhang · 2021
Cited alongside, same era.
Auctions and peer prediction for scientific peer review
Siddarth Srinivasan and Jamie Morgenstern · 2021
Cited alongside, same era.
Yuqi Pan, Zhaohua Chen, and Yuqing Kong · 2023
Later among the works it cites.
Nlp evaluation in trouble: On the need to measure llm data contamination for each benchmark
Oscar Sainz, Jon Ander Campos, Iker García-Ferrero, Julen Etxaniz, Oier Lopez de Lacalle, and Eneko Agirre · 2023
Later among the works it cites.
Can artificial intelligence help for scientific writing?
Michele Salvagno, Fabio Silvio Taccone, Alberto Giovanni Gerli, et al · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models, 2023
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, Dan Bikel, Lukas Blecher, Cristian Canton Ferrer, Moya Chen, Guillem Cucurull, David Esiobu, Jude Fernandes, Jeremy Fu, Wenyin Fu, Brian Fuller, Cynthia Gao, Vedanuj Goswami, Naman Goyal, Anthony Hartshorn, Saghar Hosseini, Rui Hou, Hakan Inan, Marcin Kardas, Viktor Kerkez, Madian Khabsa, Isabel Kloumann, Artem Korenev, Punit Singh Koura, Marie-Anne Lachaux, Thibaut Lavril, Jenya Lee, Diana Liskovich, Yinghai Lu, Yuning Mao, Xavier Martinet, Todor Mihaylov, Pushkar Mishra, Igor Molybog, Yixin Nie, Andrew Poulton, Jeremy Reizenstein, Rashi Rungta, Kalyan Saladi, Alan Schelten, Ruan Silva, Eric Michael Smith, Ranjan Subramanian, Xiaoqing Ellen Tan, Binh Tang, Ross Taylor, Adina Williams, Jian Xiang Kuan, Puxin Xu, Zheng Yan, Iliyan Zarov, Yuchen Zhang, Angela Fan, Melanie Kambadur, Sharan Narang, Aurelien Rodriguez, Robert Stojnic, Sergey Edunov, and Thomas Scialom · 2023
Later among the works it cites.
GLM-130b: An open bilingual pre-trained model
Aohan Zeng, Xiao Liu, Zhengxiao Du, Zihan Wang, Hanyu Lai, Ming Ding, Zhuoyi Yang, Yifan Xu, Wendi Zheng, Xiao Xia, Weng Lam Tam, Zixuan Ma, Yufei Xue, Jidong Zhai, Wenguang Chen, Zhiyuan Liu, Peng Zhang, Yuxiao Dong, and Jie Tang · 2023
Later among the works it cites.
Multitask peer prediction with task-dependent strategies
Yichi Zhang and Grant Schoenebeck · 2023
Later among the works it cites.
High-effort crowds: Limited liability via tournaments
Yichi Zhang and Grant Schoenebeck · 2023
Later among the works it cites.
ClusterLLM: Large language models as a guide for text clustering
Yuwei Zhang, Zihan Wang, and Jingbo Shang · 2023
Later among the works it cites.
Algorithmic robust forecast aggregation
Yongkang Guo, Jason D Hartline, Zhihuan Huang, Yuqing Kong, Anant Shah, and Fang-Yi Yu · 2024
Closest in time.
Dominantly truthful peer prediction mechanisms with a finite number of tasks
Yuqing Kong · 2024
Closest in time.
Calibrating “cheap signals” in peer review without a prior
Yuxuan Lu and Yuqing Kong · 2024
Closest in time.
Can chatgpt evaluate research quality?
Mike Thelwall · 2024
Closest in time.
Elicitationgpt: Text elicitation mechanisms via language models
Yifan Wu and Jason Hartline · 2024
Closest in time.
Spot check equivalence: an interpretable metric for information elicitation mechanisms
Shengwei Xu, Yichi Zhang, Paul Resnick, and Grant Schoenebeck · 2024
Closest in time.
Benchmarking large language models for news summarization
Tianyi Zhang, Faisal Ladhak, Esin Durmus, Percy Liang, Kathleen McKeown, and Tatsunori B Hashimoto · 2024
Closest in time.