Fetching the paper…
Reading the bibliography…
There are two main barriers to using large language models (LLMs) in clinical reasoning.
P. M. Dung, “On the acceptability of arguments and its fundamental role in nonmonotonic reasoning, logic programming and n-person games,” Artificial Intelligence , vol. 77, no. 2, pp. 321–357, 1995. [Online]. Available: https://www.sciencedirect.com/science/article/pii/000437029400041X
1995
Earlier work this paper cites.
D. Walton, Argumentation Schemes for Presumptive Reasoning , 1st ed. Routledge, 1996. [Online]. Available: https://doi.org/10.4324/9780203811160
1996
Earlier work this paper cites.
G. Klyne and J. J. Carroll. (2004) Resource description framework (rdf): Concepts and abstract syntax. W3C Recommendation. W3C. [Online]. Available: http://www.w3.org/TR/2004/REC-rdf-concepts-20040210/
2004
Earlier work this paper cites.
D. Walton, C. Reed, and F. Macagno, Argumentation Schemes , C. Reed and F. Macagno, Eds. New York: Cambridge University Press, 2008
2008
Earlier work this paper cites.
M. A. Qassas, D. Fogli, M. Giacomin, and G. Guida, “Analysis of clinical discussions based on argumentation schemes,” Procedia Computer Science , vol. 64, pp. 282–289, 2015, conference on ENTERprise Information Systems/International Conference on Project MANagement/Conference on Health and Social Care Information Systems and Technologies, CENTERIS/ProjMAN / HCist 2015 October 7-9, 2015. [Online]. Available: https://www.sciencedirect.com/science/article/pii/S1877050915026265
2015
Earlier work this paper cites.
K. Čyras, B. Delaney, D. Prociuk, F. Toni, M. Chapman, J. Domínguez, and V. Curcin, “Argumentation for explainable reasoning with conflicting medical recommendations,” CEUR Workshop Proceedings , vol. 2237, Jan. 2018, 2018 Joint Reasoning with Ambiguous and Conflicting Evidence and Recommendations in Medicine and the 3rd International Workshop on Ontology Modularity, Contextuality, and Evolution, MedRACER + WOMoCoE 2018 ; Conference date: 29-10-2018
2018
Earlier work this paper cites.
Q. Jin, B. Dhingra, Z. Liu, W. W. Cohen, and X. Lu, “Pubmedqa: A dataset for biomedical research question answering,” 2019
2019
Earlier work this paper cites.
Z. Zeng, Z. Shen, B. T. H. Tan, J. J. Chin, C. Leung, Y. Wang, Y. Chi, and C. Miao, “Explainable and Argumentation-based Decision Making with Qualitative Preferences for Diagnostics and Prognostics of Alzheimer’s Disease,” in Proceedings of the 17th International Conference on Principles of Knowledge Representation and Reasoning , 9 2020, pp. 816–826. [Online]. Available: https://doi.org/10.24963/kr.2020/84
2020
Earlier work this paper cites.
D. Jin, E. Pan, N. Oufattole, W.-H. Weng, H. Fang, and P. Szolovits, “What disease does this patient have? a large-scale open domain question answering dataset from medical exams,” 2020
2020
Earlier work this paper cites.
“ISO, IEC: AWI TS 29119-11: Software and systems engineering - software testing - part 11: Testing of AI systems,” Tech. Rep., 2020. [Online]. Available: https://www.iso.org/standard/84127.html
2020
Earlier work this paper cites.
——, “Computational argumentation and cognition,” 2021
2021
Earlier work this paper cites.
I. Sassoon, N. Kökciyan, S. Modgil, and S. Parsons, “Argumentation schemes for clinical decision support,” Argument and Computation , vol. 12, no. 3, pp. 329–355, Nov. 2021
2021
Earlier work this paper cites.
K. Singhal, S. Azizi, T. Tu, S. S. Mahdavi, J. Wei, H. W. Chung, N. Scales, A. Tanwani, H. Cole-Lewis, S. Pfohl, P. Payne, M. Seneviratne, P. Gamble, C. Kelly, N. Scharli, A. Chowdhery, P. Mansfield, B. A. y Arcas, D. Webster, G. S. Corrado, Y. Matias, K. Chou, J. Gottweis, N. Tomasev, Y. Liu, A. Rajkomar, J. Barral, C. Semturs, A. Karthikesalingam, and V. Natarajan, “Large language models encode clinical knowledge,” 2022
2022
Earlier work this paper cites.
E. Dietz, A. Kakas, and L. Michael, “Argumentation: A calculus for human-centric ai,” Frontiers in Artificial Intelligence , vol. 5, 2022. [Online]. Available: https://www.frontiersin.org/articles/10.3389/frai.2022.955579
2022
Earlier work this paper cites.
Y. Xiu, Z. Xiao, and Y. Liu, “LogicNMR: Probing the non-monotonic reasoning ability of pre-trained language models,” in Findings of the Association for Computational Linguistics: EMNLP 2022 , Y. Goldberg, Z. Kozareva, and Y. Zhang, Eds. Abu Dhabi, United Arab Emirates: Association for Computational Linguistics, Dec. 2022, pp. 3616–3626. [Online]. Available: https://aclanthology.org/2022.findings-emnlp.265
2022
Earlier work this paper cites.
A. Creswell, M. Shanahan, and I. Higgins, “Selection-inference: Exploiting large language models for interpretable logical reasoning,” 2022
2022
Cited alongside, same era.
J. Jung, L. Qin, S. Welleck, F. Brahman, C. Bhagavatula, R. L. Bras, and Y. Choi, “Maieutic prompting: Logically consistent reasoning with recursive explanations,” 2022
2022
Cited alongside, same era.
OpenAI, “Gpt-4 technical report,” 2023
2023
Cited alongside, same era.
A. Nayak, M. S. Alkaitis, K. Nayak, M. Nikolov, K. P. Weinfurt, and K. Schulman, “Comparison of history of present illness summaries generated by a chatbot and senior internal medicine residents,” JAMA Intern Med , vol. 183, no. 9, pp. 1026–1027, Sep 2023
2023
Cited alongside, same era.
H. Nori, Y. T. Lee, S. Zhang, D. Carignan, R. Edgar, N. Fusi, N. King, J. Larson, Y. Li, W. Liu, R. Luo, S. M. McKinney, R. O. Ness, H. Poon, T. Qin, N. Usuyama, C. White, and E. Horvitz, “Can generalist foundation models outcompete special-purpose tuning? case study in medicine,” 2023
2023
Later among the works it cites.
Q. Jin, Y. Yang, Q. Chen, and Z. Lu, “Genegpt: Augmenting large language models with domain tools for improved access to biomedical information,” 2023
2023
Later among the works it cites.
K. Singhal, T. Tu, J. Gottweis, R. Sayres, E. Wulczyn, L. Hou, K. Clark, S. Pfohl, H. Cole-Lewis, D. Neal, M. Schaekermann, A. Wang, M. Amin, S. Lachgar, P. Mansfield, S. Prakash, B. Green, E. Dominowska, B. A. y Arcas, N. Tomasev, Y. Liu, R. Wong, C. Semturs, S. S. Mahdavi, J. Barral, D. Webster, G. S. Corrado, Y. Matias, S. Azizi, A. Karthikesalingam, and V. Natarajan, “Towards expert-level medical question answering with large language models,” 2023
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. W. Ayers, A. Poliak, M. Dredze, E. C. Leas, Z. Zhu, J. B. Kelley, D. J. Faix, A. M. Goodman, C. A. Longhurst, M. Hogarth, and D. M. Smith, “Comparing Physician and Artificial Intelligence Chatbot Responses to Patient Questions Posted to a Public Social Media Forum,” JAMA Internal Medicine , vol. 183, no. 6, pp. 589–596, 06 2023. [Online]. Available: https://doi.org/10.1001/jamainternmed.2023.1838
2023
Cited alongside, same era.
A. Pal, L. K. Umapathi, and M. Sankarasubbu, “Med-halt: Medical domain hallucination test for large language models,” 2023
2023
Cited alongside, same era.
S. Hong, L. Xiao, and J. Chen, “An interaction model for merging multi-agent argumentation in shared clinical decision making,” in 2023 IEEE International Conference on Bioinformatics and Biomedicine (BIBM) , 2023, pp. 4304–4311
2023
Cited alongside, same era.
G. Chen, L. Cheng, L. A. Tuan, and L. Bing, “Exploring the potential of large language models in computational argumentation,” 2023
2023
Cited alongside, same era.
Y. Zhang, J. Yang, Y. Yuan, and A. C.-C. Yao, “Cumulative reasoning with large language models,” 2023
2023
Cited alongside, same era.
L. Wang, C. Ma, X. Feng, Z. Zhang, H. Yang, J. Zhang, Z. Chen, J. Tang, X. Chen, Y. Lin, W. X. Zhao, Z. Wei, and J.-R. Wen, “A survey on large language model based autonomous agents,” 2023
2023
Cited alongside, same era.
L. Pan, A. Albalak, X. Wang, and W. Y. Wang, “Logic-lm: Empowering large language models with symbolic solvers for faithful logical reasoning,” 2023
2023
Cited alongside, same era.
K. Gandhi, D. Sadigh, and N. D. Goodman, “Strategic reasoning with language models,” 2023
2023
Cited alongside, same era.
C. Li, C. Wong, S. Zhang, N. Usuyama, H. Liu, J. Yang, T. Naumann, H. Poon, and J. Gao, “Llava-med: Training a large language-and-vision assistant for biomedicine in one day,” 2023
2023
Later among the works it cites.
Q. Wu, G. Bansal, J. Zhang, Y. Wu, B. Li, E. Zhu, L. Jiang, X. Zhang, S. Zhang, J. Liu, A. H. Awadallah, R. W. White, D. Burger, and C. Wang, “Autogen: Enabling next-gen llm applications via multi-agent conversation,” 2023
2023
Later among the works it cites.
A. de Wynter and T. Yuan, “I wish to have an argument: Argumentative reasoning in large language models,” 2023
2023
Later among the works it cites.
E. Eigner and T. Händler, “Determinants of llm-assisted decision-making,” 2024
2024
Closest in time.
F. Castagna, N. Kokciyan, I. Sassoon, S. Parsons, and E. Sklar, “Computational argumentation-based chatbots: a survey,” 2024
2024
Closest in time.
J. Xie, K. Zhang, J. Chen, T. Zhu, R. Lou, Y. Tian, Y. Xiao, and Y. Su, “Travelplanner: A benchmark for real-world planning with language agents,” 2024
2024
Closest in time.
W. Shi, R. Xu, Y. Zhuang, Y. Yu, J. Zhang, H. Wu, Y. Zhu, J. Ho, C. Yang, and M. D. Wang, “Ehragent: Code empowers large language models for few-shot complex tabular reasoning on electronic health records,” 2024
2024
Closest in time.
X. Tang, A. Zou, Z. Zhang, Z. Li, Y. Zhao, X. Zhang, A. Cohan, and M. Gerstein, “Medagents: Large language models as collaborators for zero-shot medical reasoning,” 2024
2024
Closest in time.
T. Savage, A. Nayak, R. Gallo, and et al., “Diagnostic reasoning prompts reveal the potential for large language model interpretability in medicine,” npj Digital Medicine , vol. 7, p. 20, 2024. [Online]. Available: https://doi.org/10.1038/s41746-024-01010-1
2024
Closest in time.
S. Pan, L. Luo, Y. Wang, C. Chen, J. Wang, and X. Wu, “Unifying large language models and knowledge graphs: A roadmap,” IEEE Transactions on Knowledge and Data Engineering , p. 1–20, 2024. [Online]. Available: http://dx.doi.org/10.1109/TKDE.2024.3352100
2024
Closest in time.
T. Oliveira, J. Dauphin, K. Satoh, S. Tsumoto, and P. Novais, “Argumentation with goals for clinical decision support in multimorbidity,” in Proceedings of the 17th International Conference on Autonomous Agents and MultiAgent Systems , 2018, 2031–2033
2033
Closest in time.