Fetching the paper…
Reading the bibliography…
As Large Language Models (LLMs) become more prevalent, concerns about their safety, ethics, and potential biases have risen.
“The Simulation of Verbal Learning Behavior”
E.. Feigenbaum · 1963
Earlier work this paper cites.
“The Simulation of Verbal Learning Behavior”
E.. Feigenbaum · 1963
Earlier work this paper cites.
“Human Problem Solving”
A. Newell and H.. Simon · 1972
Earlier work this paper cites.
“Human Problem Solving”
A. Newell and H.. Simon · 1972
Earlier work this paper cites.
“A Computational Model of Language Acquisition in the Two-Year Old”
J… Hill · 1983
Earlier work this paper cites.
“A Computational Model of Language Acquisition in the Two-Year Old”
J… Hill · 1983
Earlier work this paper cites.
“Identifying Solution Paths in Cognitive Diagnosis”, 1985
S. Ohlsson and P. Langley · 1985
Earlier work this paper cites.
“Identifying Solution Paths in Cognitive Diagnosis”, 1985
S. Ohlsson and P. Langley · 1985
Earlier work this paper cites.
“Unsupervised Learning of Correlational Structure”
A. Chalnick and D. Billman · 1988
Earlier work this paper cites.
“Unsupervised Learning of Correlational Structure”
A. Chalnick and D. Billman · 1988
Earlier work this paper cites.
San Mateo, CA: Morgan Kaufmann, 1990
“Computational Models of Scientific Discovery and Theory Formation” · 1990
Earlier work this paper cites.
San Mateo, CA: Morgan Kaufmann, 1990
“Computational Models of Scientific Discovery and Theory Formation” · 1990
Earlier work this paper cites.
“How Real Is Fictive Motion?”, 2001
Teenie Matlock · 2001
Earlier work this paper cites.
“How Real Is Fictive Motion?”, 2001
Teenie Matlock · 2001
Earlier work this paper cites.
“A domain-specific risk-attitude scale: Measuring risk perceptions and risk behaviors”
Elke Weber, Ann-Renee Blais and Nancy Betz · 2002
Earlier work this paper cites.
“A domain-specific risk-attitude scale: Measuring risk perceptions and risk behaviors”
Elke Weber, Ann-Renee Blais and Nancy Betz · 2002
Earlier work this paper cites.
“A Domain-Specific Risk-Taking (DOSPERT) scale for adultpopulations”
Ann-Renée Blais and Elke Weber · 2006
Earlier work this paper cites.
“A Domain-Specific Risk-Taking (DOSPERT) scale for adultpopulations”
Ann-Renée Blais and Elke Weber · 2006
Earlier work this paper cites.
“Likert scale: Explored and explained”
Ankur Joshi, Saket Kale, Satish Chandel and D Pal · 2015
Earlier work this paper cites.
“Likert scale: Explored and explained”
Ankur Joshi, Saket Kale, Satish Chandel and D Pal · 2015
Earlier work this paper cites.
“Does the DOSPERT scale predict risk-taking behaviour during travel? A study using smartphones”
Andrea Farnham et al · 2018
Earlier work this paper cites.
“Glue: A multi-task benchmark and analysis platform for natural language understanding”
Alex Wang · 2018
Earlier work this paper cites.
“Unsupervised text style transfer using language models as discriminators”
Zichao Yang et al · 2018
Earlier work this paper cites.
“Does the DOSPERT scale predict risk-taking behaviour during travel? A study using smartphones”
Andrea Farnham et al · 2018
Earlier work this paper cites.
“Glue: A multi-task benchmark and analysis platform for natural language understanding”
Alex Wang · 2018
Earlier work this paper cites.
“Unsupervised text style transfer using language models as discriminators”
Zichao Yang et al · 2018
Earlier work this paper cites.
“Language (Technology) is Power: A Critical Survey of “Bias” in NLP”
Su Blodgett, Solon Barocas, Hal Daumé and Hanna Wallach · 2020
Earlier work this paper cites.
“Can GPT-3 pass a writer’s Turing test?”
Katherine Elkins and Jon Chun · 2020
Earlier work this paper cites.
“Measuring massive multitask language understanding”
Dan Hendrycks et al · 2020
Earlier work this paper cites.
“Assessing a domain-specific risk-taking construct: A meta-analysis of reliability of the DOSPERT scale”
Yiyun Shou and Joel Olney · 2020
Earlier work this paper cites.
“Language (Technology) is Power: A Critical Survey of “Bias” in NLP”
Su Blodgett, Solon Barocas, Hal Daumé and Hanna Wallach · 2020
Earlier work this paper cites.
“Can GPT-3 pass a writer’s Turing test?”
Katherine Elkins and Jon Chun · 2020
Earlier work this paper cites.
“Measuring massive multitask language understanding”
Dan Hendrycks et al · 2020
Earlier work this paper cites.
“Assessing a domain-specific risk-taking construct: A meta-analysis of reliability of the DOSPERT scale”
Yiyun Shou and Joel Olney · 2020
Earlier work this paper cites.
“Evaluating large language models trained on code”
Mark Chen et al · 2021
Cited alongside, same era.
“A survey on bias and fairness in machine learning”
Ninareh Mehrabi et al · 2021
Cited alongside, same era.
“Evaluating large language models trained on code”
Mark Chen et al · 2021
Cited alongside, same era.
“A survey on bias and fairness in machine learning”
Ninareh Mehrabi et al · 2021
Cited alongside, same era.
“Meet Your Favorite Character: Open-domain Chatbot Mimicking Fictional Characters with only a Few Utterances”
Seungju Han et al · 2022
Cited alongside, same era.
“Discovering language model behaviors with model-written evaluations”
“Stimulating the posterior parietal cortex reduces self-reported risk-taking propensity in people with tobacco use disorder”
Francesca LoFaro et al · 2024
Closest in time.
“Are Large Language Models Consistent over Value-laden Questions?”
Jared Moore, Tanvi Deshpande and Diyi Yang · 2024
Closest in time.
“GPT-4O Mini: Advancing Cost-Efficient Intelligence” Accessed: 2024-10-4, https://openai.com/index/gpt-4o-mini-advancing-cost-efficient-intelligence/
OpenAI · 2024
Closest in time.
“ProgressGym: Alignment with a Millennium of Moral Progress”
T. Qiu et al · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ethan Perez et al · 2022
Cited alongside, same era.
“Meet Your Favorite Character: Open-domain Chatbot Mimicking Fictional Characters with only a Few Utterances”
Seungju Han et al · 2022
Cited alongside, same era.
“Discovering language model behaviors with model-written evaluations”
Ethan Perez et al · 2022
Cited alongside, same era.
“Sparks of artificial general intelligence: Early experiments with gpt-4”
Sébastien Bubeck et al · 2023
Cited alongside, same era.
“Using large language models in psychology”
Dorottya Demszky et al · 2023
Cited alongside, same era.
“Toxicity in CHATGPT: Analyzing Persona-assigned Language Models”
Ameet Deshpande et al · 2023
Cited alongside, same era.
“Can AI language models replace human participants?”
Danica Dillion, Niket Tandon, Yuling Gu and Kurt Gray · 2023
Cited alongside, same era.
Yuanyi Ren et al · 2024
Closest in time.
“In-context impersonation reveals Large Language Models’ strengths and biases”
Leonard Salewski et al · 2024
Closest in time.
“CRiskEval: A Chinese Multi-Level Risk Evaluation Benchmark Dataset for Large Language Models”
Ling Shi and Deyi Xiong · 2024
Closest in time.
“Gemma 2: Improving open language models at a practical size”
Gemma Team et al · 2024
Closest in time.
“Qwen2.5: A Party of Foundation Models!” Accessed: 2024-10-4, https://qwenlm.github.io/blog/qwen2.5/ , 2024
Qwen Team · 2024
Closest in time.
Ruiyi Wang et al · 2024
Closest in time.
“Don’t Listen To Me: Understanding and Exploring Jailbreak Prompts of Large Language Models”
Z. Yu et al · 2024
Closest in time.
“R-Judge: Benchmarking Safety Risk Awareness for LLM Agents”
Tongxin Yuan et al · 2024
Closest in time.
“Expel: LLM agents are experiential learners”
Andrew Zhao et al · 2024
Closest in time.
“Claude 3.5 Sonnet” Accessed: 2024-10-4, https://www.anthropic.com/news/claude-3-5-sonnet , 2024
Anthropic · 2024
Closest in time.
“A survey on evaluation of large language models”
Yupeng Chang et al · 2024
Closest in time.
“Hypnotizability-related risky experience and behavior”
Francy Cruz-Sanabria et al · 2024
Closest in time.
“Large Language Models as Moral Experts? GPT-4 Outperforms Expert Ethicist in Providing Moral Guidance”, 2024
D. Dillion, D. Mondal, N. Tandon and K. Gray · 2024
Closest in time.
“ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools”, 2024
Team GLM, Aohan Zeng, Bin Xu and et al · 2024
Closest in time.
“Exploring the frontiers of LLMs in psychological applications: A comprehensive review”
Luoma Ke, Song Tong, Peng Chen and Kaiping Peng · 2024
Closest in time.
“The ethics of interaction: Mitigating security threats in llms”
A. Kumar, S. Singh, S.. Murty and S. Ragupathy · 2024
Closest in time.
“Quantifying ai psychology: A psychometrics benchmark for large language models”
Yuan Li et al · 2024
Closest in time.
“Stimulating the posterior parietal cortex reduces self-reported risk-taking propensity in people with tobacco use disorder”
Francesca LoFaro et al · 2024
Closest in time.
“Are Large Language Models Consistent over Value-laden Questions?”
Jared Moore, Tanvi Deshpande and Diyi Yang · 2024
Closest in time.
“GPT-4O Mini: Advancing Cost-Efficient Intelligence” Accessed: 2024-10-4, https://openai.com/index/gpt-4o-mini-advancing-cost-efficient-intelligence/
OpenAI · 2024
Closest in time.
“ProgressGym: Alignment with a Millennium of Moral Progress”
T. Qiu et al · 2024
Closest in time.
Yuanyi Ren et al · 2024
Closest in time.
“In-context impersonation reveals Large Language Models’ strengths and biases”
Leonard Salewski et al · 2024
Closest in time.
“CRiskEval: A Chinese Multi-Level Risk Evaluation Benchmark Dataset for Large Language Models”
Ling Shi and Deyi Xiong · 2024
Closest in time.
“Gemma 2: Improving open language models at a practical size”
Gemma Team et al · 2024
Closest in time.
“Qwen2.5: A Party of Foundation Models!” Accessed: 2024-10-4, https://qwenlm.github.io/blog/qwen2.5/ , 2024
Qwen Team · 2024
Closest in time.
Ruiyi Wang et al · 2024
Closest in time.
“Don’t Listen To Me: Understanding and Exploring Jailbreak Prompts of Large Language Models”
Z. Yu et al · 2024
Closest in time.
“R-Judge: Benchmarking Safety Risk Awareness for LLM Agents”
Tongxin Yuan et al · 2024
Closest in time.
“Expel: LLM agents are experiential learners”
Andrew Zhao et al · 2024
Closest in time.