Fetching the paper…
Reading the bibliography…
Recent advances in large language models (LLMs) have accelerated the development of conversational agents capable of generating human-like responses.
A coefficient of agreement for nominal scales
Cohen, J · 1960
Earlier work this paper cites.
Measuring nominal scale agreement among many raters
Fleiss, J. L · 1971
Earlier work this paper cites.
Bias, prevalence and kappa
Byrt, T., Bishop, J. & Carlin, J. B · 1993
Earlier work this paper cites.
The assessment of insight across cultures
Jacob, K. S · 2010
Earlier work this paper cites.
Handbook of Inter-Rater Reliability, 4th Edition: The Definitive Guide to Measuring The Extent of Agreement Among Raters (Advanced Analytics, LLC, 2014)
Gwet, K. L · 2014
Earlier work this paper cites.
Kaplan & Sadock’s Synopsis of Psychiatry (Lippincott Williams & Wilkins, Philadelphia, 2022), 12 ed. edn
Boland, R., Verduin, M. & MD, D. P. R · 2022
Earlier work this paper cites.
Diagnostic and Statistical Manual of Mental Disorders (5th ed., text rev.) (American Psychiatric Association Publishing, 2022)
Association, A. P · 2022
Earlier work this paper cites.
An Integrative Survey on Mental Health Conversational Agents to Bridge Computer Science and Medical Perspectives
Cho, Y. M., Rai, S., Ungar, L., Sedoc, J. & Sharath Guntuku · 2023
Earlier work this paper cites.
Systematic review and meta-analysis of AI-based conversational agents for promoting mental health and well-being
Li, H., Zhang, R., Lee, Y.-C., Kraut, R. E. & Mohr, D. C · 2023
Earlier work this paper cites.
Early detection of depression using a conversational AI bot: A non-clinical trial
Kaywan, P., Ahmed, K., Ibaida, A., Miao, Y. & Gu, B · 2023
Earlier work this paper cites.
A conversational agent framework for mental health screening: Design, implementation, and usability
Boian, R. et al · 2023
Earlier work this paper cites.
Screening for common mental health disorders: A psychometric evaluation of a chatbot system
Podina, I. R., Bucur, A.-M., Fodor, L. & Boian, R · 2023
Earlier work this paper cites.
Evaluating Conversational Agents for Mental Health: Scoping Review of Outcomes and Outcome Measurement Instruments
Jabir, A. I. et al · 2023
Earlier work this paper cites.
Prevalence of Mental Disorders and Associated Factors in Korean Adults: National Mental Health Survey of Korea 2021
Rim, S. J. et al · 2023
Earlier work this paper cites.
G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment, DOI: 10.48550/arXiv.2303.16634 (2023)
Liu, Y. et al · 2023
Cited alongside, same era.
A Survey of Large Language Models, DOI: 10.48550/arXiv.2303.18223 (2024)
Zhao, W. X. et al · 2024
Cited alongside, same era.
Large Language Models: A Survey, DOI: 10.48550/arXiv.2402.06196 (2024)
Minaee, S. et al · 2024
Cited alongside, same era.
GPT-4 Technical Report, DOI: 10.48550/arXiv.2303.08774 (2024)
OpenAI et al · 2024
Cited alongside, same era.
Large Language Model for Mental Health: A Systematic Review, DOI: 10.2196/preprints.57400 (2024)
Guo, Z. et al · 2024
Cited alongside, same era.
PsyGUARD: An Automated System for Suicide Detection and Risk Assessment in Psychological Counseling, DOI: 10.48550/arXiv.2409.20243 (2024)
Can AI Replace Human Subjects? A Large-Scale Replication of Psychological Experiments with LLMs
Cui, Z., Li, N. & Zhou, H · 2024
Later among the works it cites.
AgentClinic: A multimodal agent benchmark to evaluate AI in simulated clinical environments, DOI: 10.48550/arXiv.2405.07960 (2024)
Schmidgall, S. et al · 2024
Later among the works it cites.
ClinicalLab: Aligning Agents for Multi-Departmental Clinical Diagnostics in the Real World, DOI: 10.48550/arXiv.2406.13890 (2024)
Yan, W. et al · 2024
Later among the works it cites.
A Computational Framework for Behavioral Assessment of LLM Therapists, DOI: 10.48550/arXiv.2401.00820 (2024)
Chiu, Y. Y., Sharma, A., Lin, I. W. & Althoff, T · 2024
Later among the works it cites.
Towards a Client-Centered Assessment of LLM Therapists by Client Simulation, DOI: 10.48550/arXiv.2406.12266 (2024)
Wang, J. et al · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Qiu, H., Ma, L. & Lan, Z · 2024
Cited alongside, same era.
Conversational agents for depression screening: A systematic review
Otero-González, I., Pacheco-Lorenzo, M. R., Fernández-Iglesias, M. J. & Anido-Rifón, L. E · 2024
Cited alongside, same era.
LLM Questionnaire Completion for Automatic Psychiatric Assessment, DOI: 10.48550/arXiv.2406.06636 (2024)
Rosenman, G., Wolf, L. & Hendler, T · 2024
Cited alongside, same era.
PsycoLLM: Enhancing LLM for Psychological Understanding and Evaluation, DOI: 10.48550/arXiv.2407.05721 (2024)
Hu, J. et al · 2024
Cited alongside, same era.
Advancing Mental Health Pre-Screening: A New Custom GPT for Psychological Distress Assessment, DOI: 10.48550/arXiv.2408.01614 (2024)
Tang, J. & Shang, Y · 2024
Cited alongside, same era.
From Generation to Judgment: Opportunities and Challenges of LLM-as-a-judge, DOI: 10.48550/arXiv.2411.16594 (2024)
Li, D. et al · 2024
Cited alongside, same era.
Testing and Evaluation of Health Care Applications of Large Language Models: A Systematic Review
Bedi, S. et al · 2024
Cited alongside, same era.
Later among the works it cites.
Usage policies
OpenAI · 2024
Later among the works it cites.
Instruction Tuning for Large Language Models: A Survey, DOI: 10.48550/arXiv.2308.10792 (2024)
Zhang, S. et al · 2024
Later among the works it cites.
Worldwide Prevalence and Disability From Mental Disorders Across Childhood and Adolescence: Evidence From the Global Burden of Disease Study
Kieling, C. et al · 2024
Later among the works it cites.
LLM Leaderboard 2024
Vellum AI · 2024
Later among the works it cites.
Lost in the Middle: How Language Models Use Long Contexts
Liu, N. F. et al · 2024
Later among the works it cites.
Introducing Structured Outputs in the API
OpenAI · 2024
Later among the works it cites.
Prompt engineering
OpenAI · 2024
Later among the works it cites.
Increase output consistency (JSON mode)
Anthropic · 2024
Later among the works it cites.
Use XML tags to structure your prompts
Anthropic · 2024
Later among the works it cites.