Fetching the paper…
Reading the bibliography…
Literature research, vital for scientific work, faces the challenge of surging information volumes exceeding researchers' processing capabilities.
Lawrence S. Free online availability substantially increases a paper’s impact. Nature
2001
Earlier work this paper cites.
Mohammad S, Dorr B, Egan M, et al. Using citations to generate surveys of scientific paradigms. 2009:584-592
2009
Earlier work this paper cites.
Lok C. Speed reading: scientists are struggling to make sense of the expanding scientific literature. Corie Lok asks whether computational tools can do the hard work for them. Nature
2010
Earlier work this paper cites.
Agarwal N, Reddy RS, Kiran G, Rose C. Towards multi-document summarization of scientific articles: making interesting comparisons with SciSumm. 2011:8-15
2011
Earlier work this paper cites.
Jaidka K, Khoo C, Na J-C. Deconstructing human literature reviews–a framework for multi-document summarization. 2013:125-135
2013
Earlier work this paper cites.
Hendrycks D, Burns C, Basart S, et al. Measuring massive multitask language understanding. arXiv preprint arXiv:200903300
2020
Earlier work this paper cites.
Lewis P, Perez E, Piktus A, et al. Retrieval-augmented generation for knowledge-intensive nlp tasks. Advances in Neural Information Processing Systems
2020
Earlier work this paper cites.
Nikiforovskaya A, Kapralov N, Vlasova A, Shpynov O, Shpilman A. Automatic generation of reviews of scientific papers. IEEE; 2020:314-319
2020
Earlier work this paper cites.
Gururangan S, Marasović A, Swayamdipta S, et al. Don’t stop pretraining: Adapt language models to domains and tasks. arXiv preprint arXiv:200410964
2020
Earlier work this paper cites.
Ermel APC, Lacerda DP, Morandi MIW, Gauss L. Literature reviews: modern methods for investigating scientific and technological knowledge
2021
Earlier work this paper cites.
Wu PF, Zhang F. Recent Advances in Lead Chemisorption for Perovskite Solar Cells. Transactions of Tianjin University
2022
Earlier work this paper cites.
Ma C, Zhang WE, Guo M, Wang H, Sheng QZ. Multi-document summarization via deep learning techniques: A survey. Acm Comput Surv
2022
Earlier work this paper cites.
Kadavath S, Conerly T, Askell A, et al. Language models (mostly) know what they know. arXiv preprint arXiv:220705221
2022
Earlier work this paper cites.
Wang X, Wei J, Schuurmans D, et al. Self-consistency improves chain of thought reasoning in language models. arXiv preprint arXiv:220311171
2022
Earlier work this paper cites.
Khurana D, Koli A, Khatter K, Singh S. Natural language processing: state of the art, current trends and challenges. Multimed Tools Appl
2023
Earlier work this paper cites.
Liu Y, Han T, Ma S, et al. Summary of chatgpt-related research and perspective towards the future of large language models. Meta-Radiology
2023
Earlier work this paper cites.
Rein D, Hou BL, Stickland AC, et al. Gpqa: A graduate-level google-proof q&a benchmark. arXiv preprint arXiv:231112022
2023
Earlier work this paper cites.
White AD. The future of chemistry is language. Nature Reviews Chemistry
2023
Cited alongside, same era.
Lála J, O’Donoghue O, Shtedritski A, Cox S, Rodriques SG, White AD. Paperqa: Retrieval-augmented generative agent for scientific research. arXiv preprint arXiv:231207559
2023
Cited alongside, same era.
Wei S, Xu X, Qi X, et al. AcademicGPT: Empowering Academic Research. arXiv preprint arXiv:231112315
2023
Cited alongside, same era.
Kasanishi T, Isonuma M, Mori J, Sakata I. SciReviewGen: A Large-scale Dataset for Automatic Literature Review Generation. arXiv preprint arXiv:230515186
2023
Cited alongside, same era.
Gilardi F, Alizadeh M, Kubli M. ChatGPT outperforms crowd workers for text-annotation tasks. P Natl Acad Sci USA
2023
Cited alongside, same era.
Pu X, Gao M, Wan X. Summarization is (almost) dead. arXiv preprint arXiv:230909558
2023
Later among the works it cites.
Skarlinski MD, Cox S, Laurent JM, et al. Language agents achieve superhuman synthesis of scientific knowledge. arXiv preprint arXiv:240913740
2024
Closest in time.
Yang Z, Zhu Z. CuriousLLM: Elevating Multi-Document QA with Reasoning-Infused Knowledge Graph Prompting. arXiv preprint arXiv:240409077
2024
Closest in time.
Sami AM, Rasheed Z, Kemell K-K, et al. System for systematic literature review using multiple ai agents: Concept and an empirical evaluation. arXiv preprint arXiv:240308399
2024
Closest in time.
Agarwal S, Laradji IH, Charlin L, Pal C. LitLLM: A Toolkit for Scientific Literature Review. arXiv preprint arXiv:240201788
2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Törnberg P. Chatgpt-4 outperforms experts and crowd workers in annotating political twitter messages with zero-shot learning. arXiv preprint arXiv:230406588
2023
Cited alongside, same era.
Zheng LM, Chiang WL, Sheng Y, et al. Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena. Adv Neur In
2023
Cited alongside, same era.
Gou Z, Shao Z, Gong Y, et al. Critic: Large language models can self-correct with tool-interactive critiquing. arXiv preprint arXiv:230511738
2023
Cited alongside, same era.
Wang P, Li L, Chen L, et al. Large language models are not fair evaluators. arXiv preprint arXiv:230517926
2023
Cited alongside, same era.
Liu Y, Iter D, Xu Y, Wang S, Xu R, Zhu C. G-eval: Nlg evaluation using gpt-4 with better human alignment. arXiv preprint arXiv:230316634
2023
Cited alongside, same era.
Li Z, Wang C, Ma P, et al. Split and merge: Aligning position biases in large language model based evaluators. arXiv preprint arXiv:231001432
2023
Cited alongside, same era.
Shen C, Cheng L, Nguyen X-P, You Y, Bing L. Large language models are not yet human-level evaluators for abstractive summarization. arXiv preprint arXiv:230513091
2023
Cited alongside, same era.
Haryanto CY. LLAssist: Simple Tools for Automating Literature Review Using Large Language Models. arXiv preprint arXiv:240713993
2024
Closest in time.
Joos L, Keim DA, Fischer MT. Cutting Through the Clutter: The Potential of LLMs for Efficient Filtration in Systematic Literature Reviews. arXiv preprint arXiv:240710652
2024
Closest in time.
Li Y, Chen L, Liu A, Yu K, Wen L. ChatCite: LLM Agent with Human Workflow Guidance for Comparative Literature Summary. arXiv preprint arXiv:240302574
2024
Closest in time.
Zhang X, Li Y, Wang J, et al. Large Language Models as Evaluators for Recommendation Explanations. 2024:33-42
2024
Closest in time.
Bai Y, Ying J, Cao Y, et al. Benchmarking foundation models with language-model-as-an-examiner. Advances in Neural Information Processing Systems
2024
Closest in time.
Gao M, Hu X, Ruan J, Pu X, Wan X. Llm-based nlg evaluation: Current status and challenges. arXiv preprint arXiv:240201383
2024
Closest in time.
Lin XY, Zhen SY, Wang XH, et al. Data-Driven Design of Single-Atom Electrocatalysts with Intrinsic Descriptors for Carbon Dioxide Reduction Reaction. Transactions of Tianjin University
2024
Closest in time.
Xu Z, Jain S, Kankanhalli M. Hallucination is inevitable: An innate limitation of large language models. arXiv preprint arXiv:240111817
2024
Closest in time.
Kalai AT, Vempala SS. Calibrated language models must hallucinate. 2024:160-171
2024
Closest in time.
Cuconasu F, Trappolini G, Siciliano F, et al. The power of noise: Redefining retrieval for rag systems. 2024:719-729
2024
Closest in time.
Gupta A, Shirgaonkar A, Balaguer AdL, et al. RAG vs Fine-tuning: Pipelines, Tradeoffs, and a Case Study on Agriculture. arXiv preprint arXiv:240108406
2024
Closest in time.