Fetching the paper…
Reading the bibliography…
Knowledge discovery and collection are intelligence-intensive tasks that traditionally require significant human effort to ensure high-quality outputs.
BERT: pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Earlier work this paper cites.
Surfer100: Generating surveys from web resources, wikipedia-style
Irene Li, Alexander R. Fabbri, Rina Kawamura, Yixin Liu, Xiangru Tang, Jaesung Tae, Chang Shen, Sally Ma, Tomoe Mizutani, and Dragomir R. Radev · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever · 2021
Earlier work this paper cites.
Progressive generation of long text with pretrained language models
Bowen Tan, Zichao Yang, Maruan Al-Shedivat, Eric Xing, and Zhiting Hu · 2021
Earlier work this paper cites.
Self-rag: Learning to retrieve, generate, and critique through self-reflection
Akari Asai, Zeqiu Wu, Yizhong Wang, Avirup Sil, and Hannaneh Hajishirzi · 2023
Earlier work this paper cites.
Expository text generation: Imitate, retrieve, paraphrase
Nishant Balepur, Jie Huang, and Kevin Chen-Chuan Chang · 2023
Earlier work this paper cites.
Wikiweb2m: A page-level multimodal wikipedia dataset
Andrea Burns, Krishna Srinivasan, Joshua Ainslie, Geoff Brown, Bryan A. Plummer, Kate Saenko, Jianmo Ni, and Mandy Guo · 2023
Earlier work this paper cites.
Template-guided grammatical error feedback comment generation
Steven Coyne · 2023
Earlier work this paper cites.
Evaluating large language models on wikipedia-style survey generation
Fan Gao, Hang Jiang, Rui Yang, Qingcheng Zeng, Jinghui Lu, Moritz Blum, Dairui Liu, Tianwei She, Yuang Jiang, and Irene Li · 2023
Earlier work this paper cites.
Critic: Large language models can self-correct with tool-interactive critiquing
Zhibin Gou, Zhihong Shao, Yeyun Gong, Yelong Shen, Yujiu Yang, Nan Duan, and Weizhu Chen · 2023
Earlier work this paper cites.
Self-refine: Iterative refinement with self-feedback
Aman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan, Luyu Gao, Sarah Wiegreffe, Uri Alon, Nouha Dziri, Shrimai Prabhumoye, Yiming Yang, Sean Welleck, Bodhisattwa Prasad Majumder, Shashank Gupta, Amir Yazdanbakhsh, and Peter Clark · 2023
Earlier work this paper cites.
FActScore: Fine-grained atomic evaluation of factual precision in long form text generation
Sewon Min, Kalpesh Krishna, Xinxi Lyu, Mike Lewis, Wen-tau Yih, Pang Koh, Mohit Iyyer, Luke Zettlemoyer, and Hannaneh Hajishirzi · 2023
Earlier work this paper cites.
Refiner: Reasoning feedback on intermediate representations
Debjit Paul, Mete Ismayilzada, Maxime Peyrard, Beatriz Borges, Antoine Bosselut, Robert West, and Boi Faltings · 2023
Earlier work this paper cites.
Beyond summarization: Designing ai support for real-world expository writing tasks
Zejiang Shen, Tal August, Pao Siangliulue, Kyle Lo, Jonathan Bragg, Jeff Hammerbacher, Doug Downey, and Joseph Chee Chang · 2023
Earlier work this paper cites.
Reflexion: language agents with verbal reinforcement learning
Noah Shinn, Federico Cassano, Beck Labash, Ashwin Gopinath, Karthik Narasimhan, and Shunyu Yao · 2023
Cited alongside, same era.
Principle-driven self-alignment of language models from scratch with minimal human supervision
Zhiqing Sun, Yikang Shen, Qinhong Zhou, Hongxin Zhang, Zhenfang Chen, David D. Cox, Yiming Yang, and Chuang Gan · 2023
Cited alongside, same era.
Enhancing dialogue generation via dynamic graph knowledge aggregation
Chen Tang, Hongbo Zhang, Tyler Loakman, Chenghua Lin, and Frank Guerin · 2023
Cited alongside, same era.
Mindmap: Knowledge graph prompting sparks graph of thoughts in large language models
Yilin Wen, Zifeng Wang, and Jimeng Sun · 2023
Cited alongside, same era.
Demystifying chains, trees, and graphs of thoughts
Maciej Besta, Florim Memedi, Zhenyu Zhang, Robert Gerstenberger, Guangyuan Piao, Nils Blach, Piotr Nyczyk, Marcin Copik, Grzegorz Kwa’sniewski, Jurgen Muller, Lukas Gianinazzi, Aleš Kubček, Hubert Niewiadomski, Aidan O’Mahony, Onur Mutlu, and Torsten Hoefler · 2024
Assisting in writing wikipedia-like articles from scratch with large language models
Yijia Shao, Yucheng Jiang, Theodore A. Kanell, Peter Xu, Omar Khattab, and Monica S. Lam · 2024
Later among the works it cites.
Haochen Tan, Zhijiang Guo, Zhan Shi, Lu Xu, Zhili Liu, Yunlong Feng, Xiaoguang Li, Yasheng Wang, Lifeng Shang, Qun Liu, et al · 2024
Later among the works it cites.
Hugo Touvron et al · 2024
Later among the works it cites.
Spinning the golden thread: Benchmarking long-form generation in language models
Yuhao Wu, Ming Shan Hee, Zhiqing Hu, and Roy Ka-Wei Lee · 2024
Later among the works it cites.
Mirror: A multiple-perspective self-reflection method for knowledge-rich reasoning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
BAMBOO: A comprehensive benchmark for evaluating long text modeling capacities of large language models
Zican Dong, Tianyi Tang, Junyi Li, Wayne Xin Zhao, and Ji-Rong Wen · 2024
Cited alongside, same era.
Aaron Hurst, Adam Lerer, Adam P Goucher, Adam Perelman, Aditya Ramesh, Aidan Clark, AJ Ostrow, Akila Welihinda, Alan Hayes, Alec Radford, et al · 2024
Cited alongside, same era.
Into the unknown unknowns: Engaged human learning through participation in language model agent conversations
Yucheng Jiang, Yijia Shao, Dekun Ma, Sina J. Semnani, and Monica S. Lam · 2024
Cited alongside, same era.
DSPy: Compiling declarative language model calls into self-improving pipelines
Omar Khattab, Arnav Singhvi, Paridhi Maheshwari, Zhiyuan Zhang, Keshav Santhanam, Sri Vardhamanan, Saiful Haq, Ashutosh Sharma, Thomas T. Joshi, Hanna Moazam, Heather Miller, Matei Zaharia, and Christopher Potts · 2024
Cited alongside, same era.
Prometheus 2: An open source language model specialized in evaluating other language models
Seungone Kim, Juyoung Suk, Shayne Longpre, Bill Yuchen Lin, Jamin Shin, Sean Welleck, Graham Neubig, Moontae Lee, Kyungjae Lee, and Minjoon Seo · 2024
Cited alongside, same era.
Longlamp: A benchmark for personalized long-form text generation
Ishita Kumar, Snigdha Viswanathan, Sushrita Yerra, Alireza Salemi, Ryan A Rossi, Franck Dernoncourt, Hanieh Deilamsalehy, Xiang Chen, Ruiyi Zhang, Shubham Agarwal, et al · 2024
Cited alongside, same era.
Integrating planning into single-turn long-form text generation
Yi Liang, You Wu, Honglei Zhuang, Li Chen, Jiaming Shen, Yiling Jia, Zhen Qin, Sumit K. Sanghai, Xuanhui Wang, Carl Yang, and Michael Bendersky · 2024
Cited alongside, same era.
Hanqi Yan, Qinglin Zhu, Xinyu Wang, Lin Gui, and Yulan He · 2024
Later among the works it cites.
In-context principle learning from mistakes
Tianjun Zhang, Aman Madaan, Luyu Gao, Steven Zheng, Swaroop Mishra, Yiming Yang, Niket Tandon, and Uri Alon · 2024
Later among the works it cites.
Hypothesis generation with large language models
Yangqiaoyu Zhou, Haokun Liu, Tejes Srivastava, Hongyuan Mei, and Chenhao Tan · 2024
Later among the works it cites.
Jinze Bai et al · 2025
Closest in time.
Deepseek llm: Scaling open-source language models with longtermism
DeepSeek-AI: Xiao Bi et al · 2025
Closest in time.
Gemini: Multimodal language model
Google DeepMind · 2025
Closest in time.
Chatgpt: Optimizing language models for dialogue
OpenAI · 2025
Closest in time.
Openai o3-mini, 2025
OpenAI · 2025
Closest in time.
Omnithink: Expanding knowledge boundaries in machine writing through thinking
Zekun Xi, Wenbiao Yin, Jizhan Fang, Jialong Wu, Runnan Fang, Ningyu Zhang, Jiang Yong, Pengjun Xie, Fei Huang, and Huajun Chen · 2025
Closest in time.