Fetching the paper…
Reading the bibliography…
The advancements of large language models (LLMs) have piqued growing interest in developing LLM-based language agents to automate scientific discovery end-to-end, which has sparked both excitement and skepticism about their true capabilities.
Ohio supercomputer center, 1987
Ohio Supercomputer Center · 1987
Earlier work this paper cites.
Cellprofiler: image analysis software for identifying and quantifying cell phenotypes
Anne E. Carpenter, Thouis R. Jones, Michael R. Lamprecht, Colin Clarke, In Han Kang, Ola Friman, David A. Guertin, Joo Han Chang, Robert A. Lindquist, Jason Moffat, Polina Golland, and David M. Sabatini · 2006
Earlier work this paper cites.
Beyond the data deluge
Gordon Bell, Tony Hey, and Alex Szalay · 2009
Earlier work this paper cites.
The Fourth Paradigm: Data-Intensive Scientific Discovery
Tony Hey, Stewart Tansley, Kristin Tolle, and Jim Gray · 2009
Earlier work this paper cites.
Artificial Intelligence: A Modern Approach
Stuart Russell and Peter Norvig · 2010
Earlier work this paper cites.
Human selection of elk behavioural traits in a landscape of fear
Simone Ciuti, Tyler B. Muhlym, Dale G. Paton, Allan D. McDevitt, Marco Musiani, and Mark S. Boyce · 2012
Earlier work this paper cites.
Python materials genomics (pymatgen): A robust, open-source python library for materials analysis
Shyue Ping Ong, William Davidson Richards, Anubhav Jain, Geoffroy Hautier, Michael Kocher, Shreyas Cholia, Dan Gunter, Vincent L. Chevrier, Kristin A. Persson, and Gerbrand Ceder · 2012
Earlier work this paper cites.
Urban climate effects on extreme temperatures in madison, wisconsin, usa
Jason Schatz and Christopher J Kucharik · 2015
Earlier work this paper cites.
eofs: A library for eof analysis of meteorological, oceanographic, and climate data
Andrew Dawson · 2016
Earlier work this paper cites.
Data science: A comprehensive overview
Longbing Cao · 2017
Earlier work this paper cites.
Is multitask deep learning practical for pharma?
Bharath Ramsundar, Bowen Liu, Zhenqin Wu, Andreas Verras, Matthew Tudor, Robert P. Sheridan, and Vijay Pande · 2017
Earlier work this paper cites.
Automated Inference of Chemical Discriminants of Biological Activity , pp. 307–338
Sebastian Raschka, Anne M. Scott, Mar Huertas, Weiming Li, and Leslie A. Kuhn · 2018
Earlier work this paper cites.
Matminer: An open source toolkit for materials data mining
Logan Ward, Alexander Dunn, Alireza Faghaninia, Nils E.R. Zimmermann, Saurabh Bajaj, Qi Wang, Joseph Montoya, Jiming Chen, Kyle Bystrom, Maxwell Dylla, Kyle Chard, Mark Asta, Kristin A. Persson, G. Jeffrey Snyder, Ian Foster, and Anubhav Jain · 2018
Earlier work this paper cites.
Scanpy: large-scale single-cell gene expression data analysis
F. Alexander Wolf, Philipp Angerer, and Fabian J. Theis · 2018
Earlier work this paper cites.
Moleculenet: a benchmark for molecular machine learning
Zhenqin Wu, Bharath Ramsundar, Evan N. Feinberg, Joseph Gomes, Caleb Geniesse, Aneesh S. Pappu, Karl Leswing, and Vijay Pande · 2018
Earlier work this paper cites.
The open global glacier model (oggm) v1.1
F. Maussion, A. Butenko, N. Champollion, M. Dusch, J. Eis, K. Fourteau, P. Gregor, A. H. Jarosch, J. Landmann, F. Oesterle, B. Recinos, T. Rothenpieler, A. Vlug, C. T. Wild, and B. Marzeion · 2019
Earlier work this paper cites.
Deep Learning for the Life Sciences
Bharath Ramsundar, Peter Eastman, Patrick Walters, Vijay Pande, Karl Leswing, and Zhenqin Wu · 2019
Earlier work this paper cites.
Modeling human syllogistic reasoning:the role of “no valid conclusion”
Nicolas Riesterer, Daniel Brand, Hannah Dames, and Marco Ragni · 2019
Earlier work this paper cites.
Analyzing the differences in human reasoning viajoint nonnegative matrix factorization
Daniel Brand, Nicolas Riesterer, Hannah Dames, and Marco Ragni · 2020
Earlier work this paper cites.
Deeppurpose: A deep learning library for drug-target interaction prediction
Kexin Huang, Tianfan Fu, Lucas M Glass, Marinka Zitnik, Cao Xiao, and Jimeng Sun · 2020
Earlier work this paper cites.
The materials simulation toolkit for machine learning (mast-ml): An automated open source toolkit to accelerate data-driven materials research
Ryan Jacobs, Tam Mayeshiba, Ben Afflerbach, Luke Miles, Max Williams, Matthew Turner, Raphael Finkel, and Dane Morgan · 2020
Earlier work this paper cites.
Scirpy: a Scanpy extension for analyzing single-cell T-cell receptor-sequencing data
Gregor Sturm, Tamas Szabo, Georgios Fotakis, Marlene Haider, Dietmar Rieder, Zlatko Trajanoski, and Francesca Finotello · 2020
Earlier work this paper cites.
ProLIF: a library to encode molecular interactions as fingerprints
Cédric Bouysset and Sébastien Fiorucci · 2021
Earlier work this paper cites.
Evaluating large language models trained on code, 2021
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, Alex Ray, Raul Puri, Gretchen Krueger, Michael Petrov, Heidy Khlaaf, Girish Sastry, Pamela Mishkin, Brooke Chan, Scott Gray, Nick Ryder, Mikhail Pavlov, Alethea Power, Lukasz Kaiser, Mohammad Bavarian, Clemens Winter, Philippe Tillet, Felipe Petroski Such, Dave Cummings, Matthias Plappert, Fotios Chantzis, Elizabeth Barnes, Ariel Herbert-Voss, William Hebgen Guss, Alex Nichol, Alex Paino, Nikolas Tezak, Jie Tang, Igor Babuschkin, Suchir Balaji, Shantanu Jain, William Saunders, Christopher Hesse, Andrew N. Carr, Jan Leike, Josh Achiam, Vedant Misra, Evan Morikawa, Alec Radford, Matthew Knight, Miles Brundage, Mira Murati, Katie Mayer, Peter Welinder, Bob McGrew, Dario Amodei, Sam McCandlish, Ilya Sutskever, and Wojciech Zaremba · 2021
Earlier work this paper cites.
Accelerating high-throughput virtual screening through molecular pool-based active learning
David E. Graff, Eugene I. Shakhnovich, and Connor W. Coley · 2021
Earlier work this paper cites.
Highly accurate protein structure prediction with alphafold
John Jumper, Richard Evans, Alexander Pritzel, Tim Green, Michael Figurnov, Olaf Ronneberger, Kathryn Tunyasuvunakool, Russ Bates, Augustin Žídek, Anna Potapenko, Alex Bridgland, Clemens Meyer, Simon A. A. Kohl, Andrew J. Ballard, Andrew Cowie, Bernardino Romera-Paredes, Stanislav Nikolov, Rishub Jain, Jonas Adler, Trevor Back, Stig Petersen, David Reiman, Ellen Clancy, Michal Zielinski, Martin Steinegger, Michalina Pacholska, Tamas Berghammer, Sebastian Bodenstein, David Silver, Oriol Vinyals, Andrew W. Senior, Koray Kavukcuoglu, Pushmeet Kohli, and Demis Hassabis · 2021
Earlier work this paper cites.
Prediction and mechanistic analysis of drug-induced liver injury (dili) based on chemical structure
Anika Liu, Moritz Walter, Peter Wright, Aleksandra Bartosik, Daniela Dolciami, Abdurrahman Elbasir, Hongbin Yang, and Andreas Bender · 2021
Earlier work this paper cites.
NeuroKit2: A Python toolbox for neurophysiological signal processing
Dominique Makowski, Tam Pham, Zen J. Lau, Jan C. Brammer, François Lespinasse, Hung Pham, Christopher Schölzel, and S. H. Annabel Chen · 2021
Cited alongside, same era.
Benchmarks for interpretation of qsar models
Mariia Matveieva and Pavel Polishchuk · 2021
Cited alongside, same era.
Biopsykit: A python package for the analysis of biopsychological data
Robert Richer, Arne Küderle, Martin Ullrich, Nicolas Rohleder, and Bjoern M. Eskofier · 2021
Cited alongside, same era.
Investigating the preferences of local residents toward a proposed bus network redesign in chattanooga, tennessee
Abubakr Ziedan, Cassidy Crossland, Candace Brakewood, Philip Pugliese, and Harrison Ooi · 2021
Cited alongside, same era.
Muon: multimodal omics analysis framework
Danila Bredikhin, Ilia Kats, and Oliver Stegle · 2022
Cited alongside, same era.
A python library for probabilistic analysis of single-cell omics data
Blade: Benchmarking language model agents for data-driven science, 2024
Ken Gu, Ruoxi Shang, Ruien Jiang, Keying Kuang, Richard-John Lin, Donghe Lyu, Yue Mao, Youran Pan, Teng Wu, Jiaqian Yu, Yikun Zhang, Tianmai M. Zhang, Lanyi Zhu, Mike A. Merrill, Jeffrey Heer, and Tim Althoff · 2024
Closest in time.
WebVoyager: Building an end-to-end web agent with large multimodal models
Hongliang He, Wenlin Yao, Kaixin Ma, Wenhao Yu, Yong Dai, Hongming Zhang, Zhenzhong Lan, and Dong Yu · 2024
Closest in time.
Self-[in]correct: Llms struggle with discriminating self-generated responses, 2024
Dongwei Jiang, Jingyu Zhang, Orion Weller, Nathaniel Weir, Benjamin Van Durme, and Daniel Khashabi · 2024
Closest in time.
SWE-bench: Can language models resolve real-world github issues?
Carlos E Jimenez, John Yang, Alexander Wettig, Shunyu Yao, Kexin Pei, Ofir Press, and Karthik R Narasimhan · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Adam Gayoso, Romain Lopez, Galen Xing, Pierre Boyeau, Valeh Valiollah Pour Amiri, Justin Hong, Katherine Wu, Michael Jayasuriya, Edouard Mehlman, Maxime Langevin, Yining Liu, Jules Samaran, Gabriel Misrachi, Achille Nazaret, Oscar Clivio, Chenling Xu, Tal Ashuach, Mariano Gabitto, Mohammad Lotfollahi, Valentine Svensson, Eduardo da Veiga Beltrame, Vitalii Kleshchevnikov, Carlos Talavera-López, Lior Pachter, Fabian J. Theis, Aaron Streets, Michael I. Jordan, Jeffrey Regier, and Nir Yosef · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, brian ichter, Fei Xia, Ed Chi, Quoc V Le, and Denny Zhou · 2022
Cited alongside, same era.
Urban wildlife corridors: Building bridges for wildlife and people
Amanda J Zellmer and Barbara S Goto · 2022
Cited alongside, same era.
Autonomous chemical research with large language models
Daniil A. Boiko, Robert MacKnight, Ben Kline, and Gabe Gomes · 2023
Cited alongside, same era.
Cellprofiler: image analysis software for identifying and quantifying cell phenotypes
O. J. M. Béquignon, B. J. Bongers, W. Jespers, A. P. IJzerman, B. van der Water, and G. J. P. van Westen · 2023
Cited alongside, same era.
Mind2web: Towards a generalist agent for the web
Xiang Deng, Yu Gu, Boyuan Zheng, Shijie Chen, Samuel Stevens, Boshi Wang, Huan Sun, and Yu Su · 2023
Cited alongside, same era.
NOAA Deep Sea Corals Research and Technology Program, 1 2023
Tom Hourigan · 2023
Cited alongside, same era.
Sayash Kapoor, Benedikt Stroebl, Zachary S. Siegel, Nitya Nadgir, and Arvind Narayanan · 2024
Closest in time.
VisualWebArena: Evaluating multimodal agents on realistic visual web tasks
Jing Yu Koh, Robert Lo, Lawrence Jang, Vikram Duvvur, Ming Lim, Po-Yu Huang, Graham Neubig, Shuyan Zhou, Russ Salakhutdinov, and Daniel Fried · 2024
Closest in time.
Analyze urban heat using kriging, July 2024
Eric Krause · 2024
Closest in time.
BioMistral: A collection of open-source pretrained large language models for medical domains
Yanis Labrak, Adrien Bazoge, Emmanuel Morin, Pierre-Antoine Gourraud, Mickael Rouvier, and Richard Dufour · 2024
Closest in time.
Can gpt-4 replicate empirical software engineering research?
Jenny T. Liang, Carmen Badea, Christian Bird, Robert DeLine, Denae Ford, Nicole Forsgren, and Thomas Zimmermann · 2024
Closest in time.
Paper copilot: A self-evolving and efficient llm system for personalized academic assistance, 2024
Guanyu Lin, Tao Feng, Pengrui Han, Ge Liu, and Jiaxuan You · 2024
Closest in time.
The ai scientist: Towards fully automated open-ended scientific discovery, 2024
Chris Lu, Cong Lu, Robert Tjarko Lange, Jakob Foerster, Jeff Clune, and David Ha · 2024
Closest in time.
Large enough
MistralAI · 2024
Closest in time.
Taskbench: Benchmarking large language models for task automation
Yongliang Shen, Kaitao Song, Xu Tan, Wenqi Zhang, Kan Ren, Siyu Yuan, Weiming Lu, Dongsheng Li, and Yueting Zhuang · 2024
Closest in time.
Can llms generate novel research ideas? a large-scale human study with 100+ nlp researchers, 2024
Chenglei Si, Diyi Yang, and Tatsunori Hashimoto · 2024
Closest in time.
ADMET-AI: a machine learning ADMET platform for evaluation of large-scale chemical libraries
Kyle Swanson, Parker Walther, Jeremy Leitz, Souhrid Mukherjee, Joseph C Wu, Rabindra V Shivnaraine, and James Zou · 2024
Closest in time.
Scicode: A research coding benchmark curated by scientists, 2024
Minyang Tian, Luyu Gao, Shizhuo Dylan Zhang, Xinan Chen, Cunwei Fan, Xuefei Guo, Roland Haas, Pan Ji, Kittithat Krongchon, Yao Li, Shengyan Liu, Di Luo, Yutao Ma, Hao Tong, Kha Trinh, Chenyu Tian, Zihan Wang, Bohao Wu, Yanyu Xiong, Shengzhu Yin, Minhui Zhu, Kilian Lieret, Yanxin Lu, Genglin Liu, Yufeng Du, Tianhua Tao, Ofir Press, Jamie Callan, Eliu Huerta, and Hao Peng · 2024
Closest in time.
LLMs in the imaginarium: Tool learning through simulated trial and error
Boshi Wang, Hao Fang, Jason Eisner, Benjamin Van Durme, and Yu Su · 2024
Closest in time.
Discovery of a structural class of antibiotics with explainable deep learning
Felix Wong, Erica J. Zheng, Jacqueline A. Valeri, Nina M. Donghia, Melis N. Anahtar, Satotaka Omori, Alicia Li, Andres Cubillos-Ruiz, Aarti Krishnan, Wengong Jin, Abigail L. Manson, Jens Friedrichs, Ralf Helbig, Behnoush Hajian, Dawid K. Fiejtek, Florence F. Wagner, Holly H. Soutter, Ashlee M. Earl, Jonathan M. Stokes, Lars D. Renner, and James J. Collins · 2024
Closest in time.
Chengyue Wu, Yixiao Ge, Qiushan Guo, Jiahao Wang, Zhixuan Liang, Zeyu Lu, Ying Shan, and Ping Luo · 2024
Closest in time.
Agentless: Demystifying llm-based software engineering agents, 2024
Chunqiu Steven Xia, Yinlin Deng, Soren Dunn, and Lingming Zhang · 2024
Closest in time.
LlaSMol: Advancing large language models for chemistry with a large-scale, comprehensive, high-quality instruction tuning dataset
Botao Yu, Frazier N. Baker, Ziqi Chen, Xia Ning, and Huan Sun · 2024
Closest in time.
MAmmoTH: Building math generalist models through hybrid instruction tuning
Xiang Yue, Xingwei Qu, Ge Zhang, Yao Fu, Wenhao Huang, Huan Sun, Yu Su, and Wenhu Chen · 2024
Closest in time.
Run geoprocessing tools with python, March 2024
Paul Zandbergen · 2024
Closest in time.
Yu Zhang, Xiusi Chen, Bowen Jin, Sheng Wang, Shuiwang Ji, Wei Wang, and Jiawei Han · 2024
Closest in time.
GPT-4V(ision) is a generalist web agent, if grounded
Boyuan Zheng, Boyu Gou, Jihyung Kil, Huan Sun, and Yu Su · 2024
Closest in time.
Webarena: A realistic web environment for building autonomous agents
Shuyan Zhou, Frank F. Xu, Hao Zhu, Xuhui Zhou, Robert Lo, Abishek Sridhar, Xianyi Cheng, Tianyue Ou, Yonatan Bisk, Daniel Fried, Uri Alon, and Graham Neubig · 2024
Closest in time.