Fetching the paper…
Reading the bibliography…
Credences are mental states corresponding to degrees of confidence in propositions.
Psychological predicates
H. Putnam · 1967
Earlier work this paper cites.
The foundations of statistics
L. J. Savage · 1972
Earlier work this paper cites.
Radical interpretation
D. Lewis · 1974
Earlier work this paper cites.
The language of thought , volume 5
J. A. Fodor · 1975
Earlier work this paper cites.
Mathematics, Matter and Method: Volume 1: Philosophical Papers
H. Putnam · 1975
Earlier work this paper cites.
Representations: Philosophical essays on the foundations of cognitive science
J. A. Fodor · 1983
Earlier work this paper cites.
Psychosemantics: The problem of meaning in the philosophy of mind , volume 2
J. A. Fodor · 1987
Earlier work this paper cites.
Risk aversion as a problem of conjoint measurement
B. Hansson · 1988
Earlier work this paper cites.
A theory of content and other essays
J. A. Fodor · 1992
Earlier work this paper cites.
The failure of expected-utility theory as a theory of reason
J. Hampton · 1994
Earlier work this paper cites.
The elm and the expert: Mentalese and its semantics
J. A. Fodor · 1995
Earlier work this paper cites.
The mind doesn’t work that way: The scope and limits of computational psychology
J. A. Fodor · 2000
Earlier work this paper cites.
Theory of games and economic behavior (60th Anniversary Commemorative Edition)
J. von Neumann and O. Morgenstern · 2007
Earlier work this paper cites.
LOT 2: The language of thought revisited
J. A. Fodor · 2008
Earlier work this paper cites.
Decision theory and rationality
J. L. Bermúdez · 2009
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Imagenet large scale visual recognition challenge
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, et al · 2015
Earlier work this paper cites.
Mentalism versus behaviourism in economics: a philosophy-of-science perspective
F. Dietrich and C. List · 2016
Earlier work this paper cites.
On the interpretation of decision theory
S. Okasha · 2016
Cited alongside, same era.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Cited alongside, same era.
Glue: A multi-task benchmark and analysis platform for natural language understanding, 2019
A. Wang, A. Singh, J. Michael, F. Hill, O. Levy, and S. R. Bowman · 2019
Cited alongside, same era.
Bringing the people back in: Contesting benchmark machine learning datasets
E. Denton, A. Hanna, R. Amironesei, A. Smart, H. Nicole, and M. K. Scheuerman · 2020
Cited alongside, same era.
A general language assistant as a laboratory for alignment
A. Askell, Y. Bai, A. Chen, D. Drain, D. Ganguli, T. Henighan, A. Jones, N. Joseph, B. Mann, N. DasSarma, et al · 2021
Cited alongside, same era.
Consciousness in artificial intelligence: insights from the science of consciousness
P. Butlin, R. Long, E. Elmoznino, Y. Bengio, J. Birch, A. Constant, G. Deane, S. M. Fleming, C. Frith, X. Ji, et al · 2023
Later among the works it cites.
A survey on evaluation of large language models
Y. Chang, X. Wang, J. Wang, Y. Wu, L. Yang, K. Zhu, H. Chen, X. Yi, C. Wang, Y. Wang, et al · 2023
Later among the works it cites.
Gemini: a family of highly capable multimodal models
Gemini Team · 2023
Later among the works it cites.
A survey of language model confidence estimation and calibration
J. Geng, F. Cai, Y. Wang, H. Koeppl, P. Nakov, and I. Gurevych · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
On the dangers of stochastic parrots: Can language models be too big?
E. M. Bender, T. Gebru, A. McMillan-Major, and S. Shmitchell · 2021
Cited alongside, same era.
On the opportunities and risks of foundation models
R. Bommasani, D. A. Hudson, E. Adeli, R. Altman, S. Arora, S. von Arx, M. S. Bernstein, J. Bohg, A. Bosselut, E. Brunskill, et al · 2021
Cited alongside, same era.
Truthful ai: Developing and governing ai that does not lie
O. Evans, O. Cotton-Barratt, L. Finnveden, A. Bales, A. Balwit, P. Wills, L. Righetti, and W. Saunders · 2021
Cited alongside, same era.
Training a helpful and harmless assistant with reinforcement learning from human feedback
Y. Bai, A. Jones, K. Ndousse, A. Askell, A. Chen, N. DasSarma, D. Drain, S. Fort, D. Ganguli, T. Henighan, et al · 2022
Cited alongside, same era.
Language models (mostly) know what they know
S. Kadavath, T. Conerly, A. Askell, T. Henighan, D. Drain, E. Perez, N. Schiefer, Z. Hatfield-Dodds, N. DasSarma, E. Tran-Johnson, et al · 2022
Cited alongside, same era.
Language models show human-like content effects on reasoning
A. K. Lampinen, I. Dasgupta, S. C. Chan, A. Creswell, D. Kumaran, J. L. McClelland, and F. Hill · 2022
Cited alongside, same era.
Holistic evaluation of language models
P. Liang, R. Bommasani, T. Lee, D. Tsipras, D. Soylu, M. Yasunaga, Y. Zhang, D. Narayanan, Y. Wu, A. Kumar, et al · 2022
Cited alongside, same era.
L. Kuhn, Y. Gal, and S. Farquhar · 2023
Later among the works it cites.
Selfcheckgpt: Zero-resource black-box hallucination detection for generative large language models
P. Manakul, A. Liusie, and M. J. Gales · 2023
Later among the works it cites.
Open AI · 2023
Later among the works it cites.
Towards evaluating ai systems for moral status using self-reports
E. Perez and R. Long · 2023
Later among the works it cites.
Role play with large language models
M. Shanahan, K. McDonell, and L. Reynolds · 2023
Later among the works it cites.
K. Tian, E. Mitchell, A. Zhou, A. Sharma, R. Rafailov, H. Yao, C. Finn, and C. D. Manning · 2023
Later among the works it cites.
Sociotechnical safety evaluation of generative ai systems
L. Weidinger, M. Rauh, N. Marchal, A. Manzini, L. A. Hendricks, J. Mateos-Garcia, S. Bergman, J. Kay, C. Griffin, B. Bariach, et al · 2023
Later among the works it cites.
Can llms express their uncertainty? an empirical evaluation of confidence elicitation in llms
M. Xiong, Z. Hu, X. Lu, Y. Li, J. Fu, J. He, and B. Hooi · 2023
Later among the works it cites.
Y. Yang, E. Chern, X. Qiu, G. Neubig, and P. Liu · 2023
Later among the works it cites.
Standards for belief representations in llms
D. A. Herrmann and B. A. Levinstein · 2024
Closest in time.
Testing the general deductive reasoning capacity of large language models using ood examples
A. Saparov, R. Y. Pang, V. Padmakumar, N. Joshi, M. Kazemi, N. Kim, and H. He · 2024
Closest in time.
Evaluating the moral beliefs encoded in llms
N. Scherrer, C. Shi, A. Feder, and D. Blei · 2024
Closest in time.
Benchmarking llms via uncertainty quantification
F. Ye, M. Yang, J. Pang, L. Wang, D. F. Wong, E. Yilmaz, S. Shi, and Z. Tu · 2024
Closest in time.