Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) are at the forefront of NLP achievements but fall short in dealing with shortcut learning, factual inconsistency, and vulnerability to adversarial inputs.These shortcomings are especially critical in medical contexts, where they can misrepresent actual model capabilities.
Meta-learning surrogate models for sequential decision making
Alexandre Galashov, Jonathan Schwarz, Hyunjik Kim, Marta Garnelo, David Saxton, Pushmeet Kohli, S. M. Ali Eslami, and Yee Whye Teh. 2019 · 1903
Earlier work this paper cites.
Deep contextualized biomedical abbreviation expansion
Qiao Jin, Jinling Liu, and Xinghua Lu. 2019 · 1906
Earlier work this paper cites.
Edinburgh clinical nlp at semeval-2024 task 2: Fine-tune your model unless you have access to gpt-4
Aryo Gema, Giwon Hong, Pasquale Minervini, Luke Daines, and Beatrice Alex. 2024 · 1915
Earlier work this paper cites.
Caresai at semeval-2024 task 2: Improving natural language inference in clinical trial data using model ensemble and data explanation
Reem Abdel-Salam, Mary Adetutu Adewunmi, and Mercy Akinwale. 2024 · 1922
Earlier work this paper cites.
Bert-attack: Adversarial attack against bert using bert
Linyang Li, Ruotian Ma, Qipeng Guo, Xiangyang Xue, and Xipeng Qiu. 2020 · 2004
Earlier work this paper cites.
Language models are few-shot learners
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2005
Earlier work this paper cites.
Evidence inference 2.0: More data, better models
Jay DeYoung, Eric P. Lehman, Benjamin E. Nye, Iain James Marshall, and Byron C. Wallace. 2020 · 2005
Earlier work this paper cites.
Factors associated with participation in breast cancer treatment clinical trials
Nancy E Avis, Kevin W Smith, Carol L Link, Gabriel N Hortobagyi, and Edgardo Rivera. 2006 · 2006
Earlier work this paper cites.
Seventy-five trials and eleven systematic reviews a day: how will we ever keep up?
Hilda Bastian, Paul Glasziou, and Iain Chalmers. 2010 · 2010
Earlier work this paper cites.
Exact: automatic extraction of clinical trial characteristics from journal publications
Svetlana Kiritchenko, Berry De Bruijn, Simona Carini, Joel Martin, and Ida Sim. 2010 · 2010
Earlier work this paper cites.
Exploiting mesh indexing in medline to generate a data set for word sense disambiguation
Antonio J Jimeno-Yepes, Bridget T McInnes, and Alan R Aronson. 2011 · 2011
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Samuel R Bowman, Gabor Angeli, Christopher Potts, and Christopher D Manning. 2015 · 2015
Earlier work this paper cites.
Shortcut learning in deep neural networks
Robert Geirhos, Jörn-Henrik Jacobsen, Claudio Michaelis, Richard Zemel, Wieland Brendel, Matthias Bethge, and Felix A Wichmann. 2020 · 2020
Earlier work this paper cites.
Biobert: a pre-trained biomedical language representation model for biomedical text mining
Jinhyuk Lee, Wonjin Yoon, Sungdong Kim, Donghyeon Kim, Sunkyu Kim, Chan Ho So, and Jaewoo Kang. 2020 · 2020
Earlier work this paper cites.
Measuring and improving consistency in pretrained language models
Yanai Elazar, Nora Kassner, Shauli Ravfogel, Abhilasha Ravichander, Eduard Hovy, Hinrich Schütze, and Yoav Goldberg. 2021 · 2021
Cited alongside, same era.
A deep database of medical abbreviations and acronyms for natural language processing
Lisa Grossman Liu, Raymond H Grossman, Elliot G Mitchell, Chunhua Weng, Karthik Natarajan, George Hripcsak, and David K Vawdrey. 2021 · 2021
Cited alongside, same era.
Domain-specific language model pretraining for biomedical natural language processing
Yu Gu, Robert Tinn, Hao Cheng, Michael Lucas, Naoto Usuyama, Xiaodong Liu, Tristan Naumann, Jianfeng Gao, and Hoifung Poon. 2021 · 2021
Cited alongside, same era.
Natural language inference in context-investigating contextual reasoning over long texts
Hanmeng Liu, Leyang Cui, Jian Liu, and Yue Zhang. 2021 · 2021
Cited alongside, same era.
Palm: Scaling language modeling with pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, et al. 2022 · 2022
D-nlp at semeval-2024 task 2: Evaluating clinical inference capabilities of large language models
Duygu ALTINOK. 2024 · 2024
Closest in time.
Crcl at semeval-2024 task 2: Simple prompt optimizations
Clement Brutti-Mairesse. 2024 · 2024
Closest in time.
Rgat at semeval-2024 task 2: Biomedical natural language inference using graph attention network
Abir Chakraborty. 2024 · 2024
Closest in time.
Puer at semeval-2024 task 2: A biolinkbert approach to biomedical natural language inference
Jiaxu Dao, Zhuoying Li, Xiuzhong Tang, Xiaoli Lan, and Junde Wang. 2024 · 2024
Closest in time.
Tldr at semeval-2024 task 2: T5-generated clinical-language summaries for deberta report analysis
Spandan Das, Vinay Samuel, and Shahriar Noroozizadeh. 2024 · 2024
Closest in time.
Usmba-nlp at semeval-2024 task 2: Safe biomedical natural language inference for clinical trials using bert
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
ChatGPT and Whisper APIs
Greg Brockman, Atty Eleti, Elie Georges, Joanne Jang, Logan Kilpatrick, Rachel Lim, Luke Miller, and Michelle Pokrass. 2023 · 2023
Cited alongside, same era.
Chatgpt outperforms crowd-workers for text-annotation tasks
Fabrizio Gilardi, Meysam Alizadeh, and Maël Kubli. 2023 · 2023
Cited alongside, same era.
NLI4CT: Multi-evidence natural language inference for clinical trial reports
Mael Jullien, Marco Valentino, Hannah Frost, Paul O’Regan, Dónal Landers, and Andre Freitas. 2023a · 2023
Cited alongside, same era.
SemEval-2023 task 7: Multi-evidence natural language inference for clinical trial data
Maël Jullien, Marco Valentino, Hannah Frost, Paul O’regan, Donal Landers, and André Freitas. 2023b · 2023
Cited alongside, same era.
Saama ai research at semeval-2023 task 7: Exploring the capabilities of flan-t5 for multi-evidence natural language inference in clinical trial data
Kamal Raj Kanakarajan and Malaikannan Sankarasubbu. 2023 · 2023
Cited alongside, same era.
A symbolic framework for systematic evaluation of mathematical reasoning with transformers
Jordan Meadows, Marco Valentino, Damien Teney, and Andre Freitas. 2023 · 2023
Cited alongside, same era.
Seme at semeval-2024 task 2: Comparing masked and generative language models on natural language inference for clinical trials
Mathilde Aguiar, Pierre Zweigenbaum, and Nona Naderi. 2024 · 2024
Cited alongside, same era.
Anass Fahfouh, Abdessamad Benlahbib, Jamal Riffi, and Hamid Tairi. 2024 · 2024
Closest in time.
Lisbon computational linguists at semeval-2024 task 2: Using a mistral-7b model and data augmentation
Artur Guimarães, Bruno Martins, and João Magalhães. 2024 · 2024
Closest in time.
Saama technologies at semeval-2024 task 2: Three-module system for nli4ct enhanced by llm-generated intermediate labels
Hwanmun Kim, Kamal raj Kanakarajan, and Malaikannan Sankarasubbu. 2024 · 2024
Closest in time.
Nycu-nlp at semeval-2024 task 2: Aggregating large language models in biomedical natural language inference for clinical trials
Lung-Hao Lee, Chen-Ya Chiou, and Tzu-Mi Lin. 2024 · 2024
Closest in time.
Fzi-wim at semeval-2024 task 2: Self-consistent cot for complex nli in biomedical domain
Jin Liu and Steffen Thoma. 2024 · 2024
Closest in time.
0x.yuan at semeval-2024 task 2: Agents debating can reach consensus and produce better outcomes in medical nli task
Yu-An Lu and Hung-Yu Kao. 2024 · 2024
Closest in time.
Iitk at semeval-2024 task 2: Exploring the capabilities of llms for safe biomedical natural language inference for clinical trials
Shreyasi Mandal and Ashutosh Modi. 2024 · 2024
Closest in time.
Clac at semeval-2024 task 2: Faithful clinical trial inference
Jennifer Marks, MohammadReza Davari, and Leila Kosseim. 2024 · 2024
Closest in time.