Fetching the paper…
Reading the bibliography…
This survey examines the rapidly evolving field of Deep Research systems -- AI-powered applications that automate complex research workflows through the integration of large language models, advanced information retrieval, and autonomous reasoning capabilities.
GRADE guidelines: 3. Rating the quality of evidence
Howard Balshem, Mark Helfand, Holger J Schünemann, Andrew D Oxman, Regina Kunzand Jan Brozek, Gunn E Vist, Yngve Falck-Ytter, Joerg Meerpohl, Susan Norris, and Gordon H Guyatt. 2011 · 2011
Earlier work this paper cites.
Natural Language is a Programming Language: Applying Natural Language Processing to Software Development
Michael D. Ernst. 2017 · 2017
Earlier work this paper cites.
The System Usability Scale: Past, Present, and Future
James R. Lewis. 2018 · 2018
Earlier work this paper cites.
HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering
Zhilin Yang, Peng Qi, Saizheng Zhang, Yoshua Bengio, William W. Cohen, Ruslan Salakhutdinov, and Christopher D. Manning. 2018 · 2018
Earlier work this paper cites.
How Confident Are We About Observational Findings in Healthcare: A Benchmark Study
Martijn J. Schuemie, M. Soledad Cepeda, Marc A. Suchard, Jianxiao Yang, Yuxi Tian, Alejandro Schuler, Patrick B. Ryan, David Madigan, and George Hripcsak. 2020 · 2020
Earlier work this paper cites.
Temporal
Temporalio. 2020 · 2020
Earlier work this paper cites.
Github Copilot
Github. 2021 · 2021
Earlier work this paper cites.
BIG-bench
Google. 2021 · 2021
Earlier work this paper cites.
Learning to Assist Agents by Observing Them
Antti Keurulainen, Isak Westerlund, Samuel Kaski, and Alexander Ilin. 2021 · 2021
Earlier work this paper cites.
Explaining Reward Functions to Humans for Better Human-Robot Collaboration
Lindsay Sanneman and Julie Shah. 2021 · 2021
Earlier work this paper cites.
Accelerating science with human versus alien artificial intelligences
Jamshid Sourati and James Evans. 2021 · 2021
Earlier work this paper cites.
How Much Automation Does a Data Scientist Want?
Dakuo Wang, Q. Vera Liao, Yunfeng Zhang, Udayan Khurana, Horst Samulowitz, Soya Park, Michael Muller, and Lisa Amini. 2021 · 2021
Earlier work this paper cites.
AI Research Associate for Early-Stage Scientific Discovery
Morad Behandish, John Maxwell III, and Johan de Kleer. 2022 · 2022
Earlier work this paper cites.
Toward Building Science Discovery Machines
Abdullah Khalili and Abdelhamid Bouchachia. 2022 · 2022
Earlier work this paper cites.
TruthfulQA: Measuring How Models Mimic Human Falsehoods
Stephanie Lin, Jacob Hilton, and Owain Evans. 2022 · 2022
Earlier work this paper cites.
Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering
Pan Lu, Swaroop Mishra, Tony Xia, Liang Qiu, Kai-Wei Chang, Song-Chun Zhu, Oyvind Tafjord, Peter Clark, and Ashwin Kalyan. 2022 · 2022
Earlier work this paper cites.
Github-Copilot-Amazon-Whisperer-ChatGPT
mirayayerdem. 2022 · 2022
Earlier work this paper cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L. Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul Christiano, Jan Leike, and Ryan Lowe. 2022 · 2022
Earlier work this paper cites.
AI Assistants: A Framework for Semi-Automated Data Wrangling
Tomas Petricek, Gerrit J. J. van den Burg, Alfredo Nazábal, Taha Ceritli, Ernesto Jiménez-Ruiz, and Christopher K. I. Williams. 2022 · 2022
Earlier work this paper cites.
Effidit: Your AI Writing Assistant
Shuming Shi, Enbo Zhao, Duyu Tang, Yan Wang, Piji Li, Wei Bi, Haiyun Jiang, Guoping Huang, Leyang Cui, Xinting Huang, Cong Zhou, Yong Dai, and Dongyang Ma. 2022 · 2022
Earlier work this paper cites.
Documentation Matters: Human-Centered AI System to Assist Data Science Code Documentation in Computational Notebooks
April Yi Wang, Dakuo Wang, Jaimie Drozdal, Michael Muller, Soya Park, Justin D. Weisz, Xuye Liu, Lingfei Wu, and Casey Dugan. 2022 · 2022
Earlier work this paper cites.
Flowise: Low-code LLM Application Building Tool
Flowise AI. 2023 · 2023
Earlier work this paper cites.
Agent-based Learning of Materials Datasets from Scientific Literature
Mehrad Ansari and Seyed Mohamad Moosavi. 2023 · 2023
Earlier work this paper cites.
GPT-Researcher
assafelovic. 2023 · 2023
Earlier work this paper cites.
MegaWika: Millions of reports and their sources across 50 diverse languages
Samuel Barham, Orion Weller, Michelle Yuan, Kenton Murray, Mahsa Yarmohammadi, Zhengping Jiang, Siddharth Vashishtha, Alexander Martin, Anqi Liu, Aaron Steven White, Jordan Boyd-Graber, and Benjamin Van Durme. 2023 · 2023
Earlier work this paper cites.
ChemCrow: Augmenting large-language models with chemistry tools
Andres M Bran, Sam Cox, Oliver Schilter, Carlo Baldassari, Andrew D White, and Philippe Schwaller. 2023 · 2023
Earlier work this paper cites.
Exploring the intersection of Generative AI and Software Development
Filipe Calegario, Vanilson Burégio, Francisco Erivaldo, Daniel Moraes Costa Andrade, Kailane Felix, Nathalia Barbosa, Pedro Lucas da Silva Lucena, and César França. 2023 · 2023
Earlier work this paper cites.
The Challenges and Opportunities of AI-Assisted Writing: Developing AI Literacy for the AI Age
Peter Cardon, Carolin Fleischmann, Jolanta Aritz, Minna Logemann, and Jeanette Heidewald. 2023 · 2023
Earlier work this paper cites.
ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate
Chi-Min Chan, Weize Chen, Yusheng Su, Jianxuan Yu, Wei Xue, Shanghang Zhang, Jie Fu, and Zhiyuan Liu. 2023 · 2023
Earlier work this paper cites.
AgentVerse: Facilitating Multi-Agent Collaboration and Exploring Emergent Behaviors
Weize Chen, Yusheng Su, Jingwei Zuo, Cheng Yang, Chenfei Yuan, Chi-Min Chan, Heyang Yu, Yaxi Lu, Yi-Hsin Hung, Chen Qian, Yujia Qin, Xin Cong, Ruobing Xie, Zhiyuan Liu, Maosong Sun, and Jie Zhou. 2023 · 2023
Earlier work this paper cites.
Conversational Challenges in AI-Powered Data Science: Obstacles, Needs, and Design Opportunities
Bhavya Chopra, Ananya Singha, Anna Fariha, Sumit Gulwani, Chris Parnin, Ashish Tiwari, and Austin Z. Henley. 2023 · 2023
Earlier work this paper cites.
AutoChain
Forethought-Technologies. 2023 · 2023
Earlier work this paper cites.
AI empowering research: 10 ways how science can benefit from AI
César França. 2023 · 2023
Earlier work this paper cites.
AssistGPT: A General Multi-modal Assistant that can Plan, Execute, Inspect, and Learn
Difei Gao, Lei Ji, Luowei Zhou, Kevin Qinghong Lin, Joya Chen, Zihan Fan, and Mike Zheng Shou. 2023 · 2023
Earlier work this paper cites.
Rishab Jain and Aditya Jain. 2023 · 2023
Earlier work this paper cites.
Automated Scientific Discovery: From Equation Discovery to Autonomous Discovery Systems
Stefan Kramer, Mattia Cerrato, Sašo Džeroski, and Ross King. 2023 · 2023
Earlier work this paper cites.
Automated scholarly paper review: Concepts, technologies, and challenges
Jialiang Lin, Jiaxin Song, Zhangping Zhou, Yidong Chen, and Xiaodong Shi. 2023 · 2023
Earlier work this paper cites.
GAIA:A Benchmark for General AI Assistants
Gregoire Mialon, Clementine Fourrier, Craig Swift, Thomas Wolf, Yann LeCun, and Thomas Scialom. 2023 · 2023
Earlier work this paper cites.
How Data Scientists Review the Scholarly Literature. In Proceedings of the 2023 Conference on Human Information Interaction and Retrieval (CHIIR ’23) . ACM, 137–152
Sheshera Mysore, Mahmood Jasim, Haoru Song, Sarah Akbar, Andre Kenneth Chase Randall, and Narges Mahyar. 2023 · 2023
Earlier work this paper cites.
The Future of AI-Assisted Writing
Carlos Alves Pereira, Tanay Komarlu, and Wael Mobeirek. 2023 · 2023
Earlier work this paper cites.
Evangelos Pournaras. 2023 · 2023
Earlier work this paper cites.
AgentGPT
reworkd. 2023 · 2023
Earlier work this paper cites.
Kaushik Roy, Vedant Khandelwal, Harshul Surana, Valerie Vera, Amit Sheth, and Heather Heckman. 2023 · 2023
Earlier work this paper cites.
LlamaIndex
Run-llama. 2023 · 2023
Earlier work this paper cites.
Accelerating science with human-aware artificial intelligence
Jamshid Sourati and James Evans. 2023 · 2023
Earlier work this paper cites.
Dify: Open-source LLM Application Development Platform
Ltd. Suzhou Yuling Artificial Intelligence Technology Co. 2023 · 2023
Earlier work this paper cites.
What Should Data Science Education Do with Large Language Models?
Xinming Tu, James Zou, Weijie J. Su, and Linjun Zhang. 2023 · 2023
Earlier work this paper cites.
Thiemo Wambsganss, Xiaotian Su, Vinitra Swamy, Seyed Parsa Neshaei, Roman Rietsche, and Tanja Käser. 2023 · 2023
Earlier work this paper cites.
Self-Consistency Improves Chain of Thought Reasoning in Language Models
Xuezhi Wang, Jason Wei, Dale Schuurmans, Quoc Le, Ed Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou. 2023 · 2023
Earlier work this paper cites.
Natural Language Generation and Understanding of Big Code for AI-Assisted Programming: A Review
Man Fai Wong, Shangxin Guo, Ching Nam Hang, Siu Wai Ho, and Chee Wei Tan. 2023 · 2023
Earlier work this paper cites.
AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation
Qingyun Wu, Gagan Bansal, Jieyu Zhang, Yiran Wu, Beibin Li, Erkang Zhu, Li Jiang, Xiaoyun Zhang, Shaokun Zhang, Jiale Liu, Ahmed Hassan Awadallah, Ryen W White, Doug Burger, and Chi Wang. 2023 · 2023
Earlier work this paper cites.
Burak Yetiştiren, Işık Özsoy, Miray Ayerdem, and Eray Tüzün. 2023 · 2023
Earlier work this paper cites.
The Future of Fundamental Science Led by Generative Closed-Loop Artificial Intelligence
Hector Zenil, Jesper Tegnér, Felipe S. Abrahão, Alexander Lavin, Vipin Kumar, Jeremy G. Frey, Adrian Weller, Larisa Soldatova, Alan R. Bundy, Nicholas R. Jennings, Koichi Takahashi, Lawrence Hunter, Saso Dzeroski, Andrew Briggs, Frederick D. Gregory, Carla P. Gomes, Jon Rowe, James Evans, Hiroaki Kitano, and Ross King. 2023 · 2023
Earlier work this paper cites.
Practices and Challenges of Using GitHub Copilot: An Empirical Study. In Proceedings of the 35th International Conference on Software Engineering and Knowledge Engineering (SEKE2023, Vol. 2023) . KSI Research Inc., 124–129
Beiq Zhang, Peng Liang, Xiyu Zhou, Aakash Ahmad, and Muhammad Waseem. 2023c · 2023
Earlier work this paper cites.
AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models
Wanjun Zhong, Ruixiang Cui, Yiduo Guo, Yaobo Liang, Shuai Lu, Yanlin Wang, Amin Saied, Weizhu Chen, and Nan Duan. 2023 · 2023
Earlier work this paper cites.
CARE: Collaborative AI-Assisted Reading Environment
Dennis Zyska, Nils Dycke, Jan Buchmann, Ilia Kuznetsov, and Iryna Gurevych. 2023 · 2023
Earlier work this paper cites.
ReSearch
Agent-RL. 2024 · 2024
Earlier work this paper cites.
Beyond-RAG: Question Identification and Answer Generation in Real-Time Conversations
Garima Agrawal, Sashank Gummuluri, and Cosimo Spera. 2024 · 2024
Earlier work this paper cites.
AI achieves silver-medal standard solving International Mathematical Olympiad problems
AlphaProof and AlphaGeometry teams. 2024 · 2024
Earlier work this paper cites.
Building effective agents
Anthropic. 2024 · 2024
Earlier work this paper cites.
Model Context Protocol (MCP)
Antropic. 2024 · 2024
Earlier work this paper cites.
LLMs as Debate Partners: Utilizing Genetic Algorithms and Adversarial Search for Adaptive Arguments
Prakash Aryan. 2024 · 2024
Earlier work this paper cites.
Zahra Ashktorab, Qian Pan, Werner Geyer, Michael Desmond, Marina Danilevsky, James M. Johnson, Casey Dugan, and Michelle Bachman. 2024 · 2024
Earlier work this paper cites.
A Retrieval-Augmented Generation Framework for Academic Literature Navigation in Data Science
Ahmet Yasin Aytar, Kemal Kilic, and Kamer Kaya. 2024 · 2024
Earlier work this paper cites.
TapeAgents: a Holistic Framework for Agent Development and Optimization
Dzmitry Bahdanau, Nicolas Gontier, Gabriel Huang, Ehsan Kamalloo, Rafael Pardinas, Alex Piché, Torsten Scholak, Oleh Shliazhko, Jordan Prince Tremblay, Karam Ghanem, Soham Parikh, Mitul Tiwari, and Quaizar Vohra. 2024 · 2024
Earlier work this paper cites.
Social AI Agents Too Need to Explain Themselves
Rhea Basappa, Mustafa Tekman, Hong Lu, Benjamin Faught, Sandeep Kakar, and Ashok K. Goel. 2024 · 2024
Earlier work this paper cites.
Deceptive Patterns of Intelligent and Interactive Writing Assistants
Karim Benharrak, Tim Zindulka, and Daniel Buschek. 2024 · 2024
Earlier work this paper cites.
OceanGPT: A Large Language Model for Ocean Science Tasks
Zhen Bi, Ningyu Zhang, Yida Xue, Yixin Ou, Daxiong Ji, Guozhou Zheng, and Huajun Chen. 2024 · 2024
Earlier work this paper cites.
Drivers and Barriers of AI Adoption and Use in Scientific Research
Stefano Bianchini, Moritz Müller, and Pierre Pelletier. 2024 · 2024
Earlier work this paper cites.
Artificial Intelligence for Literature Reviews: Opportunities and Challenges
Francisco Bolanos, Angelo Salatino, Francesco Osborne, and Enrico Motta. 2024 · 2024
Earlier work this paper cites.
Exploring the Evidence-Based Beliefs and Behaviors of LLM-Based Programming Assistants
Chris Brown and Jason Cusati. 2024 · 2024
Earlier work this paper cites.
open_deep_research
btahir. 2024 · 2024
Earlier work this paper cites.
Coze Space
ByteDance. 2024 · 2024
Earlier work this paper cites.
Exploring Human-AI Collaboration in Agile: Customised LLM Meeting Assistants
Beatriz Cabrero-Daniel, Tomas Herda, Victoria Pichler, and Martin Eder. 2024 · 2024
Earlier work this paper cites.
Automated Focused Feedback Generation for Scientific Writing Assistance
Eric Chamoun, Michael Schlichktrull, and Andreas Vlachos. 2024 · 2024
Earlier work this paper cites.
Automated Code Review In Practice
Umut Cihan, Vahid Haratian, Arda İçöz, Mert Kaan Gül, Ömercan Devran, Emircan Furkan Bayendur, Baykal Mehmet Uçar, and Eray Tüzün. 2024 · 2024
Earlier work this paper cites.
The impact of AI on engineering design procedures for dynamical systems
Kristin M. de Payrebrune, Kathrin Flaßkamp, Tom Ströhla, Thomas Sattel, Dieter Bestle, Benedict Röder, Peter Eberhard, Sebastian Peitz, Marcus Stoffel, Gulakala Rutwik, Borse Aditya, Meike Wohlleben, Walter Sextro, Maximilian Raff, C. David Remy, Manish Yadav, Merten Stender, Jan van Delden, Timo Lüddecke, Sabine C. Langer, Julius Schultz, and Christopher Blech. 2024 · 2024
Earlier work this paper cites.
Multi-line AI-assisted Code Authoring
Omer Dunay, Daniel Cheng, Adam Tait, Parth Thakkar, Peter C Rigby, Andy Chiu, Imad Ahmad, Arun Ganesan, Chandra Maddila, Vijayaraghavan Murali, Ali Tayyebi, and Nachiappan Nagappan. 2024 · 2024
Earlier work this paper cites.
A Survey on RAG Meeting LLMs: Towards Retrieval-Augmented Large Language Models
Wenqi Fan, Yujuan Ding, Liangbo Ning, Shijie Wang, Hengyun Li, Dawei Yin, Tat-Seng Chua, and Qing Li. 2024 · 2024
Earlier work this paper cites.
Flowith Oracle Mode
Flowith. 2024 · 2024
Earlier work this paper cites.
Empowering Biomedical Discovery with AI Agents
Shanghua Gao, Ada Fang, Yepeng Huang, Valentina Giunchiglia, Ayush Noori, Jonathan Richard Schwarz, Yasha Ektefaie, Jovana Kondic, and Marinka Zitnik. 2024 · 2024
Earlier work this paper cites.
SciAgents: Automating scientific discovery through multi-agent intelligent graph reasoning
Alireza Ghafarollahi and Markus J. Buehler. 2024 · 2024
Earlier work this paper cites.
AutoPenBench: Benchmarking Generative Agents for Penetration Testing
Luca Gioacchini, Marco Mellia, Idilio Drago, Alexander Delsanto, Giuseppe Siracusano, and Roberto Bifulco. 2024 · 2024
Earlier work this paper cites.
Amr Gomaa, Michael Sargious, and Antonio Krüger. 2024 · 2024
Earlier work this paper cites.
Try Deep Research and our new experimental model in Gemini, your AI assistant
Google. 2024 · 2024
Earlier work this paper cites.
The Use of AI-Robotic Systems for Scientific Discovery
Alexander H. Gower, Konstantin Korovin, Daniel Brunnsåker, Filip Kronström, Gabriel K. Reder, Ievgeniia A. Tiukova, Ronald S. Reiserer, John P. Wikswo, and Ross D. King. 2024 · 2024
Cited alongside, same era.
Hilda Hadan, Derrick Wang, Reza Hadi Mogavi, Joseph Tu, Leah Zhang-Kennedy, and Lennart E. Nacke. 2024 · 2024
Cited alongside, same era.
Mining Causality: AI-Assisted Search for Instrumental Variables
Sukjin Han. 2024 · 2024
Cited alongside, same era.
AI Data Readiness Inspector (AIDRIN) for Quantitative Assessment of Data Readiness for AI. In Proceedings of the 36th International Conference on Scientific and Statistical Database Management (SSDBM 2024) . ACM, 1–12
Kaveen Hiniduma, Suren Byna, Jean Luca Bez, and Ravi Madduri. 2024 · 2024
Cited alongside, same era.
Open-operator
browserbase. 2025 · 2025
Closest in time.
agent-tars
ByteDance. 2025 · 2025
Closest in time.
open-deep-research
Nicholas Camara. 2025 · 2025
Closest in time.
EAIRA: Establishing a Methodology for Evaluating AI Models as Scientific Research Assistants
Franck Cappello, Sandeep Madireddy, Robert Underwood, Neil Getty, Nicholas Lee-Ping Chia, Nesar Ramachandra, Josh Nguyen, Murat Keceli, Tanwi Mallick, Zilinghan Li, Marieme Ngom, Chenhui Zhang, Angel Yanguas-Gil, Evan Antoniuk, Bhavya Kailkhura, Minyang Tian, Yufeng Du, Yuan-Sen Ting, Azton Wells, Bogdan Nicolae, Avinash Maurya, M. Mustafa Rafique, Eliu Huerta, Bo Li, Ian Foster, and Rick Stevens. 2025 · 2025
Closest in time.
Why Do Multi-Agent LLM Systems Fail?
Mert Cemri, Melissa Z. Pan, Shuyi Yang, Lakshya A. Agrawal, Bhavya Chopra, Rishabh Tiwari, Kurt Keutzer, Aditya Parameswaran, Dan Klein, Kannan Ramchandran, Matei Zaharia, Joseph E. Gonzalez, and Ion Stoica. 2025 · 2025
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
AiSciVision: A Framework for Specializing Large Multimodal Models in Scientific Image Classification
Brendan Hogan, Anmol Kabra, Felipe Siqueira Pacheco, Laura Greenstreet, Joshua Fan, Aaron Ferber, Marta Ummus, Alecsander Brito, Olivia Graham, Lillian Aoki, Drew Harvell, Alex Flecker, and Carla Gomes. 2024 · 2024
Cited alongside, same era.
MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework
Sirui Hong, Mingchen Zhuge, Jiaqi Chen, Xiawu Zheng, Yuheng Cheng, Ceyao Zhang, Jinlin Wang, Zili Wang, Steven Ka Shing Yau, Zijuan Lin, Liyang Zhou, Chenyu Ran, Lingfeng Xiao, Chenglin Wu, and Jürgen Schmidhuber. 2024 · 2024
Cited alongside, same era.
Auto-Deep-Research
Hong Kong University Data Science Lab. 2024 · 2024
Cited alongside, same era.
Large Language Models as Misleading Assistants in Conversation
Betty Li Hou, Kejian Shi, Jason Phang, James Aung, Steven Adler, and Rosie Campbell. 2024 · 2024
Cited alongside, same era.
Shulin Huang, Shirong Ma, Yinghui Li, Mengzuo Huang, Wuhe Zou, Weidong Zhang, and Hai-Tao Zheng. 2024 · 2024
Cited alongside, same era.
Kurando IIDA and Kenjiro MIMURA. 2024 · 2024
Cited alongside, same era.
Seyed Mohammad Ali Jafari. 2024 · 2024
Cited alongside, same era.
Disrupting Test Development with AI Assistants: Building the Base of the Test Pyramid with Three AI Coding Assistants
Vijay Joshi and Iver Band. 2024 · 2024
Cited alongside, same era.
Daniel J. H. Chung, Zhiqi Gao, Yurii Kvasiuk, Tianyi Li, Moritz Münchmeyer, Maja Rudolph, Frederic Sala, and Sai Chaitanya Tadepalli. 2025 · 2025
Closest in time.
Deep Research is now available on Gemini 2.5 Pro Experimental
Dave Citron. 2025 · 2025
Closest in time.
Devin.ai
Cognition Labs. 2025 · 2025
Closest in time.
Consensus
Consensus. 2025 · 2025
Closest in time.
Learning to Coordinate with Experts
Mohamad H. Danesh, Tu Trinh, Benjamin Plaut, and Nguyen X. Khanh. 2025 · 2025
Closest in time.
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
DeepSeek-AI, Daya Guo, Dejian Yang, Haowei Zhang, Junxiao Song, Ruoyu Zhang, Runxin Xu, Qihao Zhu, Shirong Ma, Peiyi Wang, Xiao Bi, Xiaokang Zhang, Xingkai Yu, Yu Wu, et al · 2025
Closest in time.
Akash Dhruv and Anshu Dubey. 2025 · 2025
Closest in time.
Bridging Logic Programming and Deep Learning for Explainability through ILASP
Talissa Dreossi. 2025 · 2025
Closest in time.
"It makes you think": Provocations Help Restore Critical Thinking to AI-Assisted Knowledge Work
Ian Drosos, Advait Sarkar, Xiaotong Xu, and Neil Toronto. 2025 · 2025
Closest in time.
Steffen Eger, Yong Cao, Jennifer D’Souza, Andreas Geiger, Christian Greisinger, Stephanie Gross, Yufang Hou, Brigitte Krenn, Anne Lauscher, Yizhi Li, Chenghua Lin, Nafise Sadat Moosavi, Wei Zhao, and Tristan Miller. 2025 · 2025
Closest in time.
DeepResearcher
GAIR-NLP. 2025 · 2025
Closest in time.
gemini-fullstack-langgraph-quickstart
Google. 2025 · 2025
Closest in time.
NotebookLm
Google. 2025 · 2025
Closest in time.
ChartCitor: Multi-Agent Framework for Fine-Grained Chart Visual Attribution
Kanika Goswami, Puneet Mathur, Ryan Rossi, and Franck Dernoncourt. 2025 · 2025
Closest in time.
Towards an AI co-scientist
Juraj Gottweis, Wei-Hung Weng, Alexander Daryin, Tao Tu, Anil Palepu, Petar Sirkovic, Artiom Myaskovsky, Felix Weissenberger, Keran Rong, Ryutaro Tanno, Khaled Saab, Dan Popovici, Jacob Blum, Fan Zhang, Katherine Chou, Avinatan Hassidim, Burak Gokturk, Amin Vahdat, Pushmeet Kohli, Yossi Matias, Andrew Carroll, Kavita Kulkarni, Nenad Tomasev, Vikram Dhillon, Eeshit Dhaval Vaishnav, Byron Lee, Tiago R D Costa, José R Penadés, Gary Peltz, Yunhan Xu, Annalisa Pawlosky, Alan Karthikesalingam, and Vivek Natarajan. 2025 · 2025
Closest in time.
Towards Decoding Developer Cognition in the Age of AI Assistants
Ebtesam Al Haque, Chris Brown, Thomas D. LaToza, and Brittany Johnson. 2025 · 2025
Closest in time.
Gaole He, Patrick Hemmer, Michael Vössing, Max Schemmer, and Ujwal Gadiraju. 2025 · 2025
Closest in time.
AI-Researcher
HKUDS. 2025 · 2025
Closest in time.
smolagents: open_deep_research
HuggingFace. 2025 · 2025
Closest in time.
NoTeeline: Supporting Real-Time, Personalized Notetaking with LLM-Enhanced Micronotes. In Proceedings of the 30th International Conference on Intelligent User Interfaces (IUI ’25) . ACM, 1064–1081
Faria Huq, Abdus Samee, David Chuan-En Lin, Alice Xiaodi Tang, and Jeffrey P Bigham. 2025 · 2025
Closest in time.
AIDE: AI-Driven Exploration in the Space of Code
Zhengyao Jiang, Dominik Schmidt, Dhruv Srikanth, Dixing Xu, Ian Kaplan, Deniss Jacenko, and Yuxiang Wu. 2025 · 2025
Closest in time.
node-DeepResearch
Jina AI. 2025 · 2025
Closest in time.
DSBench: How Far Are Data Science Agents from Becoming Data Science Experts?
Liqiang Jing, Zhehui Huang, Xiaoyang Wang, Wenlin Yao, Wenhao Yu, Kaixin Ma, Hongming Zhang, Xinya Du, and Dong Yu. 2025 · 2025
Closest in time.
OpenAI’s ‘deep research’ tool: is it useful for scientists?
Nicola Jones. 2025 · 2025
Closest in time.
Gemini Launches Deep Research on 2.5 Pro Aiming to Redefine AI-Powered Analysis with Strong Lead Over OpenAI
CTOL Editors Ken. 2025 · 2025
Closest in time.
Debate Helps Weak-to-Strong Generalization
Hao Lang, Fei Huang, and Yongbin Li. 2025 · 2025
Closest in time.
How to think about agent frameworks
LangChain. 2025 · 2025
Closest in time.
ChatDBG: An AI-Powered Debugging Assistant
Kyla Levin, Nicolas van Kempen, Emery D. Berger, and Stephen N. Freund. 2025 · 2025
Closest in time.
AAAR-1.0: Assessing AI’s Potential to Assist Research
Renze Lou, Hanzi Xu, Sijia Wang, Jiangshu Du, Ryo Kamoi, Xiaoxin Lu, Jian Xie, Yuxuan Sun, Yusen Zhang, Jihyun Janice Ahn, Hongchao Fang, Zhuoyang Zou, Wenchao Ma, Xi Li, Kai Zhang, Congying Xia, Lifu Huang, and Wenpeng Yin. 2025 · 2025
Closest in time.
Automated Capability Discovery via Model Self-Exploration
Cong Lu, Shengran Hu, and Jeff Clune. 2025 · 2025
Closest in time.
Generative AI Voting: Fair Collective Choice is Resilient to LLM Biases and Inconsistencies
Srijoni Majumdar, Edith Elkind, and Evangelos Pournaras. 2025 · 2025
Closest in time.
Dung Nguyen Manh, Thang Phan Chau, Nam Le Hai, Thong T. Doan, Nam V. Nguyen, Quang Pham, and Nghi D. Q. Bui. 2025 · 2025
Closest in time.
ChatGPT’s Deep Research vs. Google’s Gemini 1.5 Pro with Deep Research: A Detailed Comparison
Jonathan Mast. 2025 · 2025
Closest in time.
Gianmarco Mengaldo. 2025 · 2025
Closest in time.
MGX.dev
MGX Technologies. 2025 · 2025
Closest in time.
lightllm
ModelTC. 2025 · 2025
Closest in time.
CodeA11y: Making AI Coding Assistants Useful for Accessible Web Development
Peya Mowar, Yi-Hao Peng, Jason Wu, Aaron Steinfeld, and Jeffrey P. Bigham. 2025 · 2025
Closest in time.
From Documents to Dialogue: Building KG-RAG Enhanced AI Assistants
Manisha Mukherjee, Sungchul Kim, Xiang Chen, Dan Luo, Tong Yu, and Tung Mai. 2025 · 2025
Closest in time.
AlphaEvolve: A coding agent for scientific and algorithmic discovery
Alexander Novikov, Ngân Vũ, Marvin Eisenberger, Emilien Dupont, Po-Sen Huang, Adam Zsolt Wagner, Sergey Shirobokov, Borislav Kozlovskii, Francisco J. R. Ruiz, Abbas Mehrabian, M. Pawan Kumar, Abigail See, Swarat Chaudhuri, George Holland, Alex Davies, Sebastian Nowozin, Pushmeet Kohli, and Matej Balog. 2025 · 2025
Closest in time.
Automating Care by Self-maintainability for Full Laboratory Automation
Koji Ochiai, Yuya Tahara-Arai, Akari Kato, Kazunari Kaizu, Hirokazu Kariyazaki, Makoto Umeno, Koichi Takahashi, Genki N. Kanda, and Haruka Ozaki. 2025 · 2025
Closest in time.
OpenManus
Open Manus Team. 2025 · 2025
Closest in time.
Introducing Deep Research
OpenAI. 2025 · 2025
Closest in time.
Introducing OpenAI o3 and o4-mini
OpenAI. 2025 · 2025
Closest in time.
OpenAI Agents SDK
OpenAI. 2025 · 2025
Closest in time.
Position: AI agents should be regulated based on autonomous action sequences
Takauki Osogami. 2025 · 2025
Closest in time.
Introducing Perplexity Deep Research
Perplexity. 2025 · 2025
Closest in time.
Sonar by Perplexity
Perplexity. 2025 · 2025
Closest in time.
Long Phan, Alice Gatti, Ziwen Han, Nathaniel Li, Josephina Hu, Hugh Zhang, Chen Bo Calvin Zhang, Mohamed Shaaban, John Ling, Sean Shi, et al · 2025
Closest in time.
Benchmarking Agentic Workflow Generation
Shuofei Qiao, Runnan Fang, Zhisong Qiu, Xiaobin Wang, Ningyu Zhang, Yong Jiang, Pengjun Xie, Fei Huang, and Huajun Chen. 2025 · 2025
Closest in time.
Accelerating Scientific Research Through a Multi-LLM Framework
Joaquin Ramirez-Medina, Mohammadmehdi Ataei, and Alidad Amirfazli. 2025 · 2025
Closest in time.
Large language model for patent concept generation
Runtao Ren, Jian Ma, and Jianxi Luo. 2025 · 2025
Closest in time.
ResearchRabbit
ResearchRabbit. 2025 · 2025
Closest in time.
A Multi-Year Grey Literature Review on AI-assisted Test Automation
Filippo Ricca, Alessandro Marchetto, and Andrea Stocco. 2025 · 2025
Closest in time.
Nathalie Riche, Anna Offenwanger, Frederic Gmeiner, David Brown, Hugo Romat, Michel Pahud, Nicolai Marquardt, Kori Inkpen, and Ken Hinckley. 2025 · 2025
Closest in time.
AgentLaboratory
SamuelSchmidgall. 2025 · 2025
Closest in time.
Agent Laboratory: Using LLM Agents as Research Assistants
Samuel Schmidgall, Yusheng Su, Ze Wang, Ximeng Sun, Jialian Wu, Xiaodong Yu, Jiang Liu, Zicheng Liu, and Emad Barsoum. 2025 · 2025
Closest in time.
OpenDeepResearcher
Michael Shumer. 2025 · 2025
Closest in time.
Welcome to the Era of Experience
David Silver and Richard Sutton. 2025 · 2025
Closest in time.
FutureHouse Platform: Superintelligent AI Agents for Scientific Discovery
Michael Skarlinski, Tyler Nadolski, James Braza, Remo Storni, Mayk Caldas, Ludovico Mitchener, Michaela Hinks, Andrew White, and Sam Rodriques. 2025 · 2025
Closest in time.
Performance Evaluation of Large Language Models in Statistical Programming
Xinyi Song, Kexin Xie, Lina Lee, Ruizhe Chen, Jared M. Clark, Hao He, Haoran He, Jie Min, Xinlei Zhang, Simin Zheng, Zhiyang Zhang, Xinwei Deng, and Yili Hong. 2025 · 2025
Closest in time.
Haoyang Su, Renqi Chen, Shixiang Tang, Zhenfei Yin, Xinzhe Zheng, Jinzhe Li, Biqing Qi, Qi Wu, Hui Li, Wanli Ouyang, Philip Torr, Bowen Zhou, and Nanqing Dong. 2025 · 2025
Closest in time.
AutoAgent: A Fully-Automated and Zero-Code Framework for LLM Agents
Jiabin Tang, Tianyu Fan, and Chao Huang. 2025 · 2025
Closest in time.
deep_research_agent
Yan Tang. 2025 · 2025
Closest in time.
Thanh-Dat Truong, Hoang-Quan Nguyen, Xuan-Bac Nguyen, Ashley Dowling, Xin Li, and Khoa Luu. 2025 · 2025
Closest in time.
Enabling AI Scientists to Recognize Innovation: A Domain-Agnostic Algorithm for Assessing Novelty
Yao Wang, Mingxuan Cui, and Arthur Jiang. 2025 · 2025
Closest in time.
AI’s deep research revolution: Transforming biomedical literature analysis
Ying-Mei Wang and Tzeng-J Chen. 2025 · 2025
Closest in time.
AI Benchmark Deep Dive: Gemini 2.5 and Humanity’s Last Exam
Sarah Welsh. 2025 · 2025
Closest in time.
CycleResearcher: Improving Automated Research via Automated Review
Yixuan Weng, Minjun Zhu, Guangsheng Bao, Hongbo Zhang, Jindong Wang, Yue Zhang, and Linyi Yang. 2025 · 2025
Closest in time.
Agentic Reasoning: Reasoning LLMs with Tools for the Deep Research
Junde Wu, Jiayuan Zhu, and Yuyuan Liu. 2025 · 2025
Closest in time.
Grok 3 Beta - The age of reasoning agents
xAI. 2025 · 2025
Closest in time.
Minerva: A Programmable Memory Test Benchmark for Language Models
Menglin Xia, Victor Ruehle, Saravan Rajmohan, and Reza Shokri. 2025 · 2025
Closest in time.
UGPhysics: A Comprehensive Benchmark for Undergraduate Physics Reasoning with Large Language Models
Xin Xu, Qiyun Xu, Tong Xiao, Tianhao Chen, Yuchen Yan, Jiaxin Zhang, Shizhe Diao, Can Yang, and Yang Wang. 2025 · 2025
Closest in time.
Knowledge Retrieval Based on Generative AI
Te-Lun Yang, Jyi-Shane Liu, Yuen-Hsien Tseng, and Jyh-Shing Roger Jang. 2025 · 2025
Closest in time.
Knowledge Synthesis of Photosynthesis Research Using a Large Language Model
Seungri Yoon, Woosang Jeon, Sanghyeok Choi, Taehyeong Kim, and Tae In Ahn. 2025 · 2025
Closest in time.
Dolphin: Moving Towards Closed-loop Auto-research through Thinking, Practice, and Feedback
Jiakang Yuan, Xiangchao Yan, Shiyang Feng, Bo Zhang, Tao Chen, Botian Shi, Wanli Ouyang, Yu Qiao, Lei Bai, and Bowen Zhou. 2025 · 2025
Closest in time.
deep-research
David Zhang. 2025 · 2025
Closest in time.
Raigul Zheldibayeva. 2025 · 2025
Closest in time.
AutoGLM-Research
Zhipu AI. 2025 · 2025
Closest in time.
DeepReview: Improving LLM-based Paper Review with Human-like Deep Thinking Process
Minjun Zhu, Yixuan Weng, Linyi Yang, and Yue Zhang. 2025 · 2025
Closest in time.