Fetching the paper…
Reading the bibliography…
Open large language models (LLMs) have significantly advanced the field of natural language processing, showcasing impressive performance across various tasks.Despite the significant advancements in LLMs, their effective operation still relies heavily on human input to accurately guide the dialogue flow, with agent tuning being a crucial optimization technique that involves human adjustments to the model for better response to such guidance.Addressing this dependency, our work introduces the TinyAgent model, trained on a meticulously curated high-quality dataset.
Short-term memory and sentence processing: Evidence from neuropsychology
R. Martin. 1993 · 1993
Earlier work this paper cites.
Scaling laws for neural language models
Jared Kaplan, Sam McCandlish, Tom Henighan, Tom B Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeffrey Wu, and Dario Amodei. 2020 · 2001
Earlier work this paper cites.
The piazza peer data management system
A. Halevy, Z. Ives, J. Madhavan, P. Mork, Dan Suciu, and I. Tatarinov. 2004 · 2004
Earlier work this paper cites.
Selecting informative contexts improves language model fine-tuning
Richard J. Antonello, Javier Turek, and Alexander G. Huth. 2020 · 2005
Earlier work this paper cites.
The demise of short-term memory revisited: empirical and computational investigations of recency effects
E. Davelaar, Y. Goshen-Gottstein, Amir Ashkenazi, H. Haarmann, and M. Usher. 2005 · 2005
Earlier work this paper cites.
Comparison of approaches to service deployment
V. Talwar, Qinyi Wu, C. Pu, W. Yan, G. Jung, and D. Milojicic. 2005 · 2005
Earlier work this paper cites.
A collaborative e-learning system based on multi-agent
Jun Wang, Yong-Hong Sun, Z. Fan, and Yan Liu. 2005 · 2005
Earlier work this paper cites.
Kai Zhao, Yongduan Song, CL Philip Chen, and Long Chen. 2021 · 2005
Earlier work this paper cites.
Learning to summarize from human feedback
Nisan Stiennon, Long Ouyang, Jeff Wu, Daniel M. Ziegler, Ryan J. Lowe, Chelsea Voss, Alec Radford, Dario Amodei, and Paul Christiano. 2020 · 2009
Earlier work this paper cites.
Linguistic relativity revisited: The interaction between l1 and l2 in thinking, learning, and production
Hye K Pae et al. 2012 · 2012
Earlier work this paper cites.
Automatic image dataset construction from click-through logs using deep neural network
Yalong Bai, Kuiyuan Yang, Wei Yu, Chang Xu, Wei-Ying Ma, and T. Zhao. 2015 · 2015
Earlier work this paper cites.
Periodic event-triggered synchronization of linear multi-agent systems with communication delays
Eloy García, Yongcan Cao, and D. Casbeer. 2015 · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, K. Kavukcuoglu, David Silver, Andrei A. Rusu, J. Veness, Marc G. Bellemare, A. Graves, Martin A. Riedmiller, A. Fidjeland, Georg Ostrovski, Stig Petersen, Charlie Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, D. Kumaran, Daan Wierstra, S. Legg, and D. Hassabis. 2015 · 2015
Earlier work this paper cites.
Multiagent systems in construction: A ten-year review
Xin Liang, G. Shen, and Shanshan Bu. 2016 · 2016
Earlier work this paper cites.
A new ai evaluation cosmos: Ready to play the game?
José Hernández-Orallo, Marco Baroni, Jordi Bieger, Nader Chmait, David L Dowe, Katja Hofmann, Fernando Martínez-Plumed, Claes Strannegård, and Kristinn R Thórisson. 2017 · 2017
Earlier work this paper cites.
Deal or no deal? end-to-end learning for negotiation dialogues
Mike Lewis, Denis Yarats, Yann N Dauphin, Devi Parikh, and Dhruv Batra. 2017 · 2017
Earlier work this paper cites.
Apprentice: Using knowledge distillation techniques to improve low-precision network accuracy
Asit K. Mishra and Debbie Marr. 2017 · 2017
Earlier work this paper cites.
Mastering the game of go without human knowledge
David Silver, Julian Schrittwieser, Karen Simonyan, Ioannis Antonoglou, Aja Huang, Arthur Guez, Thomas Hubert, Lucas Baker, Matthew Lai, Adrian Bolton, et al. 2017 · 2017
Earlier work this paper cites.
A novel consensus algorithm for second‐order multi‐agent systems without velocity measurements
Wentao Zhang, Yang Liu, Jianquan Lu, and Jinde Cao. 2017 · 2017
Earlier work this paper cites.
Broad learning system: An effective and efficient incremental learning system without the need for deep architecture
C. L. P. Chen and Zhulin Liu. 2018 · 2018
Earlier work this paper cites.
Cloudmf: Model-driven management of multi-cloud applications
N. Ferry, F. Chauvel, Hui Song, A. Rossini, Maksym Lushpenko, and Arnor Solberg. 2018 · 2018
Earlier work this paper cites.
Universal language model fine-tuning for text classification
Jeremy Howard and Sebastian Ruder. 2018 · 2018
Cited alongside, same era.
Cross-modal attentional context learning for rgb-d object detection
Guanbin Li, Yukang Gan, Hejun Wu, Nong Xiao, and Liang Lin. 2018 · 2018
Cited alongside, same era.
Know what you don’t know: Unanswerable questions for squad
Pranav Rajpurkar, Robin Jia, and Percy Liang. 2018 · 2018
Cited alongside, same era.
Actor-critic for multi-agent system with variable quantity of agents
Guihong Wang and Jinglun Shi. 2019 · 2018
Cited alongside, same era.
Bidirectional lstm with attention mechanism and convolutional layer for text classification
Gang Liu and Jiabao Guo. 2019 · 2019
Cited alongside, same era.
Cooperative tuning of multi-agent optimal control systems
Zehui Lu, Wanxin Jin, S. Mou, and B. Anderson. 2022 · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L. Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, J. Schulman, Jacob Hilton, Fraser Kelton, Luke E. Miller, Maddie Simens, Amanda Askell, P. Welinder, P. Christiano, J. Leike, and Ryan J. Lowe. 2022 · 2022
Later among the works it cites.
Webshop: Towards scalable real-world web interaction with grounded language agents
Shunyu Yao, Howard Chen, John Yang, and Karthik Narasimhan. 2022 · 2022
Later among the works it cites.
Opt: Open pre-trained transformer language models
Susan Zhang, Stephen Roller, Naman Goyal, Mikel Artetxe, Moya Chen, Shuohui Chen, Christopher Dewan, Mona Diab, Xian Li, Xi Victoria Lin, et al. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
José R Vázquez-Canteli and Zoltán Nagy. 2019 · 2019
Cited alongside, same era.
Coordinated behavior of cooperative agents using deep reinforcement learning
Elhadji Amadou Oury Diallo, Ayumi Sugiyama, and T. Sugawara. 2020 · 2020
Cited alongside, same era.
Knowledge distillation: A survey
Jianping Gou, B. Yu, S. Maybank, and D. Tao. 2020 · 2020
Cited alongside, same era.
Blockchain-based multiparty computation system
Kai Lu and Chongyang Zhang. 2020 · 2020
Cited alongside, same era.
Understanding and improving continuous integration and delivery practice using data from the wild
M. Di Penta. 2020 · 2020
Cited alongside, same era.
Large language models associate muslims with violence
Abubakar Abid, Maheen Farooqi, and James Zou. 2021 · 2021
Cited alongside, same era.
A decentralized federated learning framework via committee mechanism with convergence guarantee
Chunjiang Che, Xiaoli Li, Chuan Chen, Xiaoyu He, and Zibin Zheng. 2021 · 2021
Cited alongside, same era.
Jinze Bai, Shuai Bai, Yunfei Chu, Zeyu Cui, Kai Dang, Xiaodong Deng, Yang Fan, Wenbin Ge, Yu Han, Fei Huang, et al. 2023 · 2023
Later among the works it cites.
Pangu-agent: A fine-tunable generalist agent with structured reasoning
Filippos Christianos, Georgios Papoudakis, Matthieu Zimmer, Thomas Coste, Zhihao Wu, Jingxuan Chen, Khyati Khandelwal, James Doran, Xidong Feng, Jiacheng Liu, et al. 2023 · 2023
Later among the works it cites.
Emergent cooperation and strategy adaptation in multi-agent systems: An extended coevolutionary theory with llms
I de Zarzà, J de Curtò, Gemma Roig, Pietro Manzoni, and Carlos T Calafate. 2023 · 2023
Later among the works it cites.
Camel: Communicative agents for” mind” exploration of large scale language model society
Guohao Li, Hasan Abed Al Kader Hammoud, Hani Itani, Dmitrii Khizbullin, and Bernard Ghanem. 2023 · 2023
Later among the works it cites.
Agentbench: Evaluating llms as agents
Xiao Liu, Hao Yu, Hanchen Zhang, Yifan Xu, Xuanyu Lei, Hanyu Lai, Yu Gu, Hangliang Ding, Kaiwen Men, Kejuan Yang, et al. 2023 · 2023
Later among the works it cites.
Gpt-4 technical report
OpenAI. 2023 · 2023
Later among the works it cites.
Steven I. Ross, Fernando Martinez, Stephanie Houde, Michael J. Muller, and Justin D. Weisz. 2023 · 2023
Later among the works it cites.
Code llama: Open foundation models for code
Baptiste Roziere, Jonas Gehring, Fabian Gloeckle, Sten Sootla, Itai Gat, Xiaoqing Ellen Tan, Yossi Adi, Jingyu Liu, Tal Remez, Jérémy Rapin, et al. 2023 · 2023
Later among the works it cites.
Noveen Sachdeva and Julian McAuley. 2023 · 2023
Later among the works it cites.
Multi-agent collaboration: Harnessing the power of intelligent llm agents
Yashar Talebirad and Amirhossein Nadiri. 2023 · 2023
Later among the works it cites.
Rolecraft-glm: Advancing personalized role-playing in large language models
Meiling Tao, Xuechen Liang, Tianyu Shi, Lei Yu, and Yiting Xie. 2023 · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al. 2023 · 2023
Later among the works it cites.
Zephyr: Direct distillation of lm alignment
Lewis Tunstall, Edward Beeching, Nathan Lambert, Nazneen Rajani, Kashif Rasul, Younes Belkada, Shengyi Huang, Leandro von Werra, Clémentine Fourrier, Nathan Habib, et al. 2023 · 2023
Later among the works it cites.
Baichuan 2: Open large-scale language models
Aiyuan Yang, Bin Xiao, Bingning Wang, Borong Zhang, Ce Bian, Chao Yin, Chenxu Lv, Da Pan, Dian Wang, Dong Yan, et al. 2023 · 2023
Later among the works it cites.
Agenttuning: Enabling generalized agent abilities for llms
Aohan Zeng, Mingdao Liu, Rui Lu, Bowen Wang, Xiao Liu, Yuxiao Dong, and Jie Tang. 2023 · 2023
Later among the works it cites.
Tinyllama: An open-source small language model
Peiyuan Zhang, Guangtao Zeng, Tianduo Wang, and Wei Lu. 2024 · 2024
Closest in time.