Fetching the paper…
Reading the bibliography…
Despite the vast repository of global medical knowledge predominantly being in English, local languages are crucial for delivering tailored healthcare services, particularly in areas with limited medical resources.
Head-qa: A healthcare dataset for complex reasoning
David Vilares and Carlos Gómez-Rodríguez. 2019 · 1906
Earlier work this paper cites.
Pubmedqa: A dataset for biomedical research question answering
Qiao Jin, Bhuwan Dhingra, Zhengping Liu, William W Cohen, and Xinghua Lu. 2019 · 1909
Earlier work this paper cites.
Qinghaosu (artemisinin): an antimalarial drug from china
Daniel L Klayman. 1985 · 1985
Earlier work this paper cites.
Towards a multilingual medical lexicon
Kornél Markó, Robert Baud, Pierre Zweigenbaum, Lars Borin, Magnus Merkel, and Stefan Schulz. 2006 · 2006
Earlier work this paper cites.
Di Jin, Eileen Pan, Nassim Oufattole, Wei-Hung Weng, Hanyi Fang, and Peter Szolovits. 2020 · 2009
Earlier work this paper cites.
Usage of multilingual mobile translation applications in clinical settings
Urs-Vito Albrecht, Marianne Behrends, Herbert K Matthies, Ute von Jan, et al. 2013 · 2013
Earlier work this paper cites.
Improving medical communication: skills for a complex (and multilingual) clinical world
Peter G Brindley, Katherine E Smith, Pierre Cardinal, Francois LeBlanc, et al. 2014 · 2014
Earlier work this paper cites.
Adaptation of machine translation for multilingual information retrieval in the medical domain
Pavel Pecina, Ondřej Dušek, Lorraine Goeuriot, Jan Hajič, Jaroslava Hlaváčová, Gareth JF Jones, Liadh Kelly, Johannes Leveling, David Mareček, Michal Novák, et al. 2014 · 2014
Earlier work this paper cites.
Determinants of prakriti, the human constitution types of indian traditional medicine and its correlation with contemporary science
Harish Rotti, Ritu Raval, Suchitra Anchan, Ravishankara Bellampalli, Sameer Bhale, Ramachandra Bharadwaj, Balakrishna K Bhat, Amrish P Dedge, Vikram Ram Dhumal, GG Gangadharan, et al. 2014 · 2014
Earlier work this paper cites.
The traditional medicine and modern medicine from natural products
Haidan Yuan, Qianqian Ma, Li Ye, and Guangchun Piao. 2016 · 2016
Earlier work this paper cites.
Think you have solved question answering? try arc, the ai2 reasoning challenge
Peter Clark, Isaac Cowhey, Oren Etzioni, Tushar Khot, Ashish Sabharwal, Carissa Schoenick, and Oyvind Tafjord. 2018 · 2018
Earlier work this paper cites.
Clear-simple corpus for medical french
Natalia Grabar and Rémi Cardon. 2018 · 2018
Earlier work this paper cites.
Named entity recognition in hindi using hyperspace analogue to language and conditional random field
Arti Jain and Anuja Arora. 2018 · 2018
Earlier work this paper cites.
Medical exam question answering with large-scale reading comprehension
Xiao Zhang, Ji Wu, Zhiyang He, Xien Liu, and Ying Su. 2018 · 2018
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2019 · 2019
Earlier work this paper cites.
The pile: An 800gb dataset of diverse text for language modeling
Leo Gao, Stella Biderman, Sid Black, Laurence Golding, Travis Hoppe, Charles Foster, Jason Phang, Horace He, Anish Thite, Noa Nabeshima, et al. 2020 · 2020
Cited alongside, same era.
Measuring massive multitask language understanding
Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt. 2020 · 2020
Cited alongside, same era.
Group differences between countries and between languages in pain-related beliefs, coping, and catastrophizing in chronic pain: a systematic review
Saurab Sharma, Alexandra Ferreira-Valente, Amanda C de C. Williams, J Haxby Abbott, José Pais-Ribeiro, and Mark P Jensen. 2020 · 2020
Cited alongside, same era.
Spanish biomedical crawled corpus: A large, diverse dataset for spanish biomedical language models
Casimiro Pio Carrino, Jordi Armengol-Estapé, Ona de Gibert Bonet, Asier Gutiérrez-Fandiño, Aitor Gonzalez-Agirre, Martin Krallinger, and Marta Villegas. 2021 · 2021
Cited alongside, same era.
Amplify-instruct: Synthetically generated diverse multi-turn conversations for efficient llm training
Luigi Daniele and Suphavadeeprasit. 2023 · 2023
Later among the works it cites.
Freedomintelligence sharegpt-language
FreedomIntelligence. 2023 · 2023
Later among the works it cites.
https://huggingface.co/datasets/krisfu/awesome-llm-datasets-only-Chinese/tree/main/sft-phase-processed
krisfu. 2023 · 2023
Later among the works it cites.
Fast inference from transformers via speculative decoding
Yaniv Leviathan, Matan Kalman, and Yossi Matias. 2023 · 2023
Later among the works it cites.
Towards expert-level medical question answering with large language models
Karan Singhal, Tao Tu, Juraj Gottweis, Rory Sayres, Ellery Wulczyn, Le Hou, Kevin Clark, Stephen Pfohl, Heather Cole-Lewis, Darlene Neal, et al. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Multilingual consultations in urgent medical care
Antoon Cox and Katrijn Maryns. 2021 · 2021
Cited alongside, same era.
Overview of bioasq 2021-mesinesp track. evaluation of advance hierarchical classification techniques for scientific literature, patents and clinical trials
Luis Gasco, Anastasios Nentidis, Anastasia Krithara, Darryl Estrada-Zavala, Renato Toshiyuki Murasaki, Elena Primo-Peña, Cristina Bojo Canales, Georgios Paliouras, Martin Krallinger, et al. 2021 · 2021
Cited alongside, same era.
Dexperts: Decoding-time controlled text generation with experts and anti-experts
Alisa Liu, Maarten Sap, Ximing Lu, Swabha Swayamdipta, Chandra Bhagavatula, Noah A Smith, and Yejin Choi. 2021 · 2021
Cited alongside, same era.
MAQA: Medical Arabic Q&A Dataset
Mohammed Abdelhay and Ammar Mohammed. 2022 · 2022
Cited alongside, same era.
Medmcqa: A large-scale multi-subject multi-choice dataset for medical domain question answering
Ankit Pal, Logesh Kumar Umapathi, and Malaikannan Sankarasubbu. 2022 · 2022
Cited alongside, same era.
WuDaoCorpora Text
Zhao Xue, Hanyu Zhao, Sha Yuan, and Yequan Wang. 2022 · 2022
Cited alongside, same era.
Zhengyun Zhao, Qiao Jin, Fangyuan Chen, Tuorui Peng, and Sheng Yu. 2022 · 2022
Cited alongside, same era.
Jinze Bai, Shuai Bai, Yunfei Chu, Zeyu Cui, Kai Dang, Xiaodong Deng, Yang Fan, Wenbin Ge, Yu Han, Fei Huang, et al. 2023 · 2023
Cited alongside, same era.
Alpaca: A strong, replicable instruction-following model
Rohan Taori, Ishaan Gulrajani, Tianyi Zhang, Yann Dubois, Xuechen Li, Carlos Guestrin, Percy Liang, and Tatsunori B Hashimoto. 2023 · 2023
Later among the works it cites.
https://huggingface.co/datasets/Vezora/Tested-22k-Python-Alpaca
Vezora. 2023 · 2023
Later among the works it cites.
Pmc-llama: Further finetuning llama on medical papers
Chaoyi Wu, Xiaoman Zhang, Ya Zhang, Yanfeng Wang, and Weidi Xie. 2023 · 2023
Later among the works it cites.
Wizardlm: Empowering large language models to follow complex instructions
Can Xu, Qingfeng Sun, Kai Zheng, Xiubo Geng, Pu Zhao, Jiazhan Feng, Chongyang Tao, and Daxin Jiang. 2023 · 2023
Later among the works it cites.
Mammoth: Building math generalist models through hybrid instruction tuning
Xiang Yue, Xingwei Qu, Ge Zhang, Yao Fu, Wenhao Huang, Huan Sun, Yu Su, and Wenhu Chen. 2023 · 2023
Later among the works it cites.
Huatuogpt, towards taming language model to be a doctor
Hongbo Zhang, Junying Chen, Feng Jiang, Fei Yu, Zhihong Chen, Jianquan Li, Guiming Chen, Xiangbo Wu, Zhiyi Zhang, Qingying Xiao, et al. 2023 · 2023
Later among the works it cites.
Abhimanyu Dubey et al. 2024 · 2024
Closest in time.
Biomistral: A collection of open-source pretrained large language models for medical domains
Yanis Labrak, Adrien Bazoge, Emmanuel Morin, Pierre-Antoine Gourraud, Mickael Rouvier, and Richard Dufour. 2024 · 2024
Closest in time.
Towards building multilingual language model for medicine
Pengcheng Qiu, Chaoyi Wu, Xiaoman Zhang, Weixiong Lin, Haicheng Wang, Ya Zhang, Yanfeng Wang, and Weidi Xie. 2024 · 2024
Closest in time.