Fetching the paper…
Reading the bibliography…
Polis is a platform that leverages machine intelligence to scale up deliberative processes.
On lines and planes of closest fit to systems of points in space
K. Pearson · 1901
Earlier work this paper cites.
Attention is not Explanation, May 2019
S. Jain and B. C. Wallace · 1902
Earlier work this paper cites.
Evaluating the Underlying Gender Bias in Contextualized Word Embeddings, Apr. 2019
C. Basta, M. R. Costa-jussà, and N. Casas · 1904
Earlier work this paper cites.
Measuring Bias in Contextualized Word Representations, June 2019
K. Kurita, N. Vyas, A. Pareek, A. W. Black, and Y. Tsvetkov · 1906
Earlier work this paper cites.
Attention is not not Explanation, Sept. 2019
S. Wiegreffe and Y. Pinter · 1908
Earlier work this paper cites.
Fine-Tuning Language Models from Human Preferences, Jan. 2020
D. M. Ziegler, N. Stiennon, J. Wu, T. B. Brown, A. Radford, D. Amodei, P. Christiano, and G. Irving · 1909
Earlier work this paper cites.
Social Bias Frames: Reasoning about Social and Power Implications of Language, Apr. 2020
M. Sap, S. Gabriel, L. Qin, D. Jurafsky, N. A. Smith, and Y. Choi · 1911
Earlier work this paper cites.
Hate speech detection: Challenges and solutions
S. MacAvaney, H.-R. Yao, E. Yang, K. Russell, N. Goharian, and O. Frieder · 1932
Earlier work this paper cites.
The Structural Transformation of the Public Sphere
J. Habermas · 1962
Earlier work this paper cites.
Some methods for classification and analysis of multivariate observations
J. B. MacQueen · 1967
Earlier work this paper cites.
A Unifying Tool for Linear Multivariate Statistical Methods: The RV- Coefficient
P. Robert and Y. Escoufier · 1976
Earlier work this paper cites.
The Theory of Communicative Action
J. Habermas · 1981
Earlier work this paper cites.
The Science Question in Feminism
S. Harding · 1986
Earlier work this paper cites.
Active learning with statistical models
D. A. Cohn, Z. Ghahramani, and M. I. Jordan · 1996
Earlier work this paper cites.
Reinforcement Learning: An Introduction
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
QuickCheck: a lightweight tool for random testing of Haskell programs
K. Claessen and J. Hughes · 2000
Earlier work this paper cites.
Hierarchical topic models and the nested chinese restaurant process
T. Griffiths, M. Jordan, J. Tenenbaum, and D. Blei · 2003
Earlier work this paper cites.
Don’t Think of an Elephant!: Know Your Values and Frame the Debate–The Essential Guide for Progressives
G. Lakoff, H. Dean, and D. Hazen · 2004
Earlier work this paper cites.
Infinite latent feature models and the Indian buffet process
Z. Ghahramani and T. L. Griffiths · 2005
Earlier work this paper cites.
Nonviolent communication : a language of life
M. B. Rosenberg · 2005
Earlier work this paper cites.
Gaussian processes for machine learning
C. E. Rasmussen · 2006
Earlier work this paper cites.
Randoop: feedback-directed random testing for Java
C. Pacheco and M. D. Ernst · 2007
Earlier work this paper cites.
The Political Mind: A Cognitive Scientist’s Guide to Your Brain and Its Politics
G. Lakoff · 2009
Earlier work this paper cites.
Machine learning: a probabilistic perspective
K. P. Murphy · 2012
Earlier work this paper cites.
Wikum: Bridging Discussion Forums and Wikis Using Recursive Summarization
A. X. Zhang, L. Verou, and D. Karger · 2017
Earlier work this paper cites.
Townhall meeting in Kentucky turns tables on polarization, 2023
E. Barry · 2018
Earlier work this paper cites.
The simple but ingenious system Taiwan uses to crowdsource its laws
C. Horton · 2018
Cited alongside, same era.
Testing Tech for Consensus in a Purple Town | Civicist, Mar. 2018
J. McKenzie · 2018
Cited alongside, same era.
First-ever civic assembly gives residents chance to be heard
D. Sergent · 2018
Cited alongside, same era.
Grounding interactive machine learning tool design in how non-experts actually build models
Q. Yang, J. Suh, N.-C. Chen, and G. Ramos · 2018
Cited alongside, same era.
Faster Peace via Inclusivity: An Efficient Paradigm to Understand Populations in Conflict Zones
J. Bilich, M. Varga, D. Masood, and A. Konya · 2019
Cited alongside, same era.
Building Consensus and Compromise on Uber in Taiwan, Sept. 2019
CPI · 2019
Cited alongside, same era.
Detecting Hate Speech with GPT-3, Mar. 2022
K.-L. Chiu, A. Collins, and R. Alexander · 2022
Later among the works it cites.
Spurious correlations in reference-free evaluation of text generation
E. Durmus, F. Ladhak, and T. Hashimoto · 2022
Later among the works it cites.
Predictability and Surprise in Large Generative Models
D. Ganguli, D. Hernandez, L. Lovitt, N. DasSarma, T. Henighan, A. Jones, N. Joseph, J. Kernion, B. Mann, A. Askell, Y. Bai, A. Chen, T. Conerly, D. Drain, N. Elhage, S. E. Showk, S. Fort, Z. Hatfield-Dodds, S. Johnston, S. Kravec, N. Nanda, K. Ndousse, C. Olsson, D. Amodei, D. Amodei, T. Brown, J. Kaplan, S. McCandlish, C. Olah, and J. Clark · 2022
Later among the works it cites.
Elicitation Inference Optimization for Multi-Principal-Agent Alignment
A. Konya, Y. L. Qiu, M. P. Varga, and A. Ovadya · 2022
Later among the works it cites.
Illustrating Reinforcement Learning from Human Feedback (RLHF)
N. Lambert, L. Castricato, L. von Werra, and A. Havrilla · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Language models are few-shot learners
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, and A. Askell · 2020
Cited alongside, same era.
FEQA: A question answering evaluation framework for faithfulness assessment in abstractive summarization
E. Durmus, H. He, and M. Diab · 2020
Cited alongside, same era.
Social Biases in NLP Models as Barriers for Persons with Disabilities
B. Hutchinson, V. Prabhakaran, E. Denton, K. Webster, Y. Zhong, and S. Denuyl · 2020
Cited alongside, same era.
Evaluating the factual consistency of abstractive text summarization
W. Kryscinski, B. McCann, C. Xiong, and R. Socher · 2020
Cited alongside, same era.
On faithfulness and factuality in abstractive summarization
J. Maynez, S. Narayan, B. Bohnet, and R. McDonald · 2020
Cited alongside, same era.
Learning to summarize with human feedback
N. Stiennon, L. Ouyang, J. Wu, D. Ziegler, R. Lowe, C. Voss, A. Radford, D. Amodei, and P. F. Christiano · 2020
Cited alongside, same era.
Later among the works it cites.
Open Democracy: Reinventing Popular Rule for the Twenty-First Century
H. Landemore · 2022
Later among the works it cites.
Evaluating human-language model interaction, 2022
M. Lee, M. Srivastava, A. Hardy, J. Thickstun, E. Durmus, A. Paranjape, I. Gerard-Ursin, X. L. Li, F. Ladhak, F. Rong, R. E. Wang, M. Kwon, J. S. Park, H. Cao, T. Lee, R. Bommasani, M. Bernstein, and P. Liang · 2022
Later among the works it cites.
Probabilistic Machine Learning: An introduction
K. P. Murphy · 2022
Later among the works it cites.
Social simulacra: Creating populated prototypes for social computing systems
J. S. Park, L. Popowski, C. J. Cai, M. R. Morris, P. Liang, and M. S. Bernstein · 2022
Later among the works it cites.
Ignore Previous Prompt: Attack Techniques For Language Models, Nov. 2022
F. Perez and I. Ribeiro · 2022
Later among the works it cites.
ChatGPT: Optimizing Language Models for Dialogue, Nov. 2022
J. Schulman, B. Zoph, C. Kim, J. Hilton, J. Menick, J. Weng, J. F. Ceron Uribe, L. Fedus, L. Metz, M. Pokorny, R. Gontijo Lopes, S. Zhao, A. Vijayvergiya, E. Sigler, A. Perelman, C. Voss, M. Heaton, J. Parish, D. Cummings, R. Nayak, V. Balcom, D. Schnurr, T. Kaftan, C. Hallacy, N. Turley, N. Deutsch, V. Goel, J. Ward, A. Konstantinidis, W. Zaremba, L. Ouyang, L. Bogdonoff, J. Gross, D. Medina, S. Yoo, T. Lee, R. Lowe, D. Mossing, J. Huizinga, R. Jiang, C. Wainwright, D. Almeida, S. Lin, M. Zhang, K. Xiao, K. Slama, S. Bills, A. Gray, J. Leike, J. Pachocki, P. Tillet, S. Jain, G. Brockman, N. Ryder, A. Paino, Q. Yuan, C. Winter, B. Wang, M. Bavarian, I. Babuschkin, S. Sidor, I. Kanitscheider, M. Pavlov, M. Plappert, N. Tezak, H. Jun, W. Zhuk, V. Pong, L. Kaiser, J. Tworek, A. Carr, L. Weng, S. Agarwal, K. Cobbe, V. Kosaraju, A. Power, S. Polu, J. Han, R. Puri, S. Jain, B. Chess, C. Gibson, O. Boiko, E. Parparita, A. Tootoonchian, K. Kosic, and C. Hesse · 2022
Later among the works it cites.
Google plans giant AI language model supporting world’s 1,000 most spoken languages, Nov. 2022
J. Vincent · 2022
Later among the works it cites.
Birdwatch: Crowd Wisdom and Bridging Algorithms can Inform Understanding and Reduce the Spread of Misinformation, Oct. 2022
S. Wojcik, S. Hilgard, N. Judd, D. Mocanu, S. Ragain, K. Coleman, M. B. F. Hunzaker, and J. Baxter · 2022
Later among the works it cites.
URL https://www.societylibrary.org
The Society Library, Mar. 2023 · 2023
Closest in time.
Marked personas: Using natural language prompts to measure stereotypes in language models, 2023
M. Cheng, E. Durmus, and D. Jurafsky · 2023
Closest in time.
Language Models Trained on Media Diets Can Predict Public Opinion, Mar. 2023
E. Chu, J. Andreas, S. Ansolabehere, and D. Roy · 2023
Closest in time.
Opinion | Can A.I. and Democracy Fix Each Other?
P. Coy · 2023
Closest in time.
Romania PM unveils AI ‘adviser’ to tell him what people think in real time
A. France-Presse · 2023
Closest in time.
K. Greshake, S. Abdelnabi, S. Mishra, C. Endres, T. Holz, and M. Fritz · 2023
Closest in time.
The political ideology of conversational AI: Converging evidence on ChatGPT’s pro-environmental, left-libertarian orientation, Jan. 2023
J. Hartmann, J. Schwenzow, and M. Witte · 2023
Closest in time.
Probabilistic Machine Learning: Advanced Topics
K. P. Murphy · 2023
Closest in time.
GPT-4 Technical Report, Mar. 2023
OpenAI · 2023
Closest in time.
Whose Opinions Do Language Models Reflect?, Mar. 2023
S. Santurkar, E. Durmus, F. Ladhak, C. Lee, P. Liang, and T. Hashimoto · 2023
Closest in time.
Explaining Patterns in Data with Language Models via Interpretable Autoprompting, Jan. 2023
C. Singh, J. X. Morris, J. Aneja, A. M. Rush, and J. Gao · 2023
Closest in time.
LLaMA: Open and Efficient Foundation Language Models, Feb. 2023
H. Touvron, T. Lavril, G. Izacard, X. Martinet, M.-A. Lachaux, T. Lacroix, B. Rozière, N. Goyal, E. Hambro, F. Azhar, A. Rodriguez, A. Joulin, E. Grave, and G. Lample · 2023
Closest in time.
Benchmarking Large Language Models for News Summarization, Jan. 2023
T. Zhang, F. Ladhak, E. Durmus, P. Liang, K. McKeown, and T. B. Hashimoto · 2023
Closest in time.