Fetching the paper…
Reading the bibliography…
The performance of Large Language Models (LLMs) is fundamentally determined by the contextual information provided during inference.
An architectural style for self-adaptive multi-agent systems, arXiv preprint arXiv:1909.03475, 2019
Danny Weyns and F. Oquendo · 1909
Earlier work this paper cites.
On formally undecidable propositions of principia mathematica and related systems
K. Gödel, B. Meltzer, and R. Schlegel · 1966
Earlier work this paper cites.
Context-dependent memory in two natural environments: on land and underwater
D. Godden and A. Baddeley · 1975
Earlier work this paper cites.
A robust layered control system for a mobile robot
Rodney A. Brooks · 1986
Earlier work this paper cites.
Prediction error-driven memory consolidation for continual learning: On the case of adaptive greenhouse models
Guido Schillaci, Uwe Schmidt, and Luis Miranda · 1987
Earlier work this paper cites.
A model for interference and forgetting
G. M. Mensink and J. Raaijmakers · 1988
Earlier work this paper cites.
Kqml-a language and protocol for knowledge and information exchange
Tim Finin, Richard Fritzson, Donald P McKay, Robin McEntire, et al · 1994
Earlier work this paper cites.
Why there are complementary learning systems in the hippocampus and neocortex: insights from the successes and failures of connectionist models of learning and memory
James L. McClelland, B. McNaughton, and R. O’Reilly · 1995
Earlier work this paper cites.
Act-r: A theory of higher level cognition and its relation to visual attention
John R. Anderson, M. Matessa, and C. Lebiere · 1997
Earlier work this paper cites.
Desire: Modelling multi-agent systems in a compositional formal framework
F. Brazier, B. Dunin-Keplicz, N. Jennings, and Jan Treur · 1997
Earlier work this paper cites.
Language-dependent recall of autobiographical memories
Viorica Marian and U. Neisser · 2000
Earlier work this paper cites.
Yara Rizk, Abhishek Bhandwalder, S. Boag, Tathagata Chakraborti, Vatche Isahagian, Y. Khazaeni, Falk Pollock, and Merve Unuvar · 2001
Earlier work this paper cites.
A structured language model based on context-sensitive probabilistic left-corner parsing
D. H. V. Uytsel, Filip Van Aelten, and Dirk Van Compernolle · 2001
Earlier work this paper cites.
A distributed representation of temporal context
Marc W Howard and M. Kahana · 2002
Earlier work this paper cites.
Memory for multidimensional source information
T. Meiser and A. Bröder · 2002
Earlier work this paper cites.
A stochastic parser based on an slm with arboreal context trees
Shinsuke Mori · 2002
Earlier work this paper cites.
The relationship between episodic memory and context: clues from memory errors made while under stress
L. Nadel, Jessica D. Payne, and W. J. Jacobs · 2002
Earlier work this paper cites.
Towards a logic of rational agency
W. Hoek and M. Wooldridge · 2003
Earlier work this paper cites.
Motivation-dependent responses in the human caudate nucleus
Mauricio R. Delgado, V. Stenger, and J. Fiez · 2004
Earlier work this paper cites.
Multi-speaker language modeling
Gang Ji and J. Bilmes · 2004
Earlier work this paper cites.
Language models are few-shot learners, arXiv preprint arXiv:2005.14165, 20202
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei · 2005
Earlier work this paper cites.
Hongchao Fang, Sicheng Wang, Meng Zhou, Jiayuan Ding, and Pengtao Xie · 2005
Earlier work this paper cites.
Managing goals and resources in dynamic environments
E. Gordon and B. Logan · 2005
Earlier work this paper cites.
Context based text-generation using lstm networks, arXiv preprint arXiv:2005.00048, 2020
S. Santhanam · 2005
Earlier work this paper cites.
Multi-agent system concepts theory and application phases
Adel Al-Jumaily · 2006
Earlier work this paper cites.
Semantic web technology for agent communication protocols
Idoia Berges, J. Bermúdez, A. Goñi, and A. Illarramendi · 2008
Earlier work this paper cites.
Aligning agent communication protocols - a pragmatic approach
Maricela Claudia Bravo and Martha Coronel · 2008
Earlier work this paper cites.
Complex and adaptive dynamical systems, arXiv preprint arXiv:0807.4838, 2008
C. Gros · 2008
Earlier work this paper cites.
Computational logic foundations of kgp agents
A. Kakas, P. Mancarella, F. Sadri, Kostas Stathis, and Francesca Toni · 2008
Earlier work this paper cites.
Automaton in or out: run-time plan optimization for xml stream processing
Hong Su, Elke A. Rundensteiner, and Murali Mani · 2008
Earlier work this paper cites.
The acquisition of robust and flexible cognitive skills
N. Taatgen, David Huss, D. Dickison, and John R. Anderson · 2008
Earlier work this paper cites.
Call graph profiling for multi agent systems
Dinh Doan Van Bien, David Lillis, and Rem W. Collier · 2009
Earlier work this paper cites.
Multi-agent systems in control engineering: a survey
Fatemeh Daneshfar and H. Bevrani · 2009
Earlier work this paper cites.
D. Ghica · 2009
Earlier work this paper cites.
Nithin Holla, Pushkar Mishra, H. Yannakoudakis, and Ekaterina Shutova · 2009
Earlier work this paper cites.
A context maintenance and retrieval model of organizational processes in free recall
Sean M. Polyn, K. Norman, and M. Kahana · 2009
Earlier work this paper cites.
Flexible memory networks
C. Curto, A. Degeratu, and V. Itskov · 2010
Earlier work this paper cites.
Toward a unified catalog of implemented cognitive architectures
A. Samsonovich · 2010
Earlier work this paper cites.
Oscillatory patterns in temporal lobe reveal context reinstatement during memory search
Jeremy R. Manning, Sean M. Polyn, G. Baltuch, B. Litt, and M. Kahana · 2011
Earlier work this paper cites.
Game-theoretic multiagent reinforcement learning, arXiv preprint arXiv:2011.00583, 2020
Yaodong Yang, Chengdong Ma, Zihan Ding, S. McAleer, Chi Jin, and Jun Wang · 2011
Earlier work this paper cites.
An overview of recent progress in the study of distributed multi-agent coordination
Yongcan Cao, Wenwu Yu, W. Ren, and Guanrong Chen · 2012
Earlier work this paper cites.
Component & service-based agent systems: Self-osgi
Mauro Dragone · 2012
Earlier work this paper cites.
Neural changes underlying the development of episodic memory during middle childhood
S. Ghetti and S. Bunge · 2012
Earlier work this paper cites.
Optogenetic stimulation of a hippocampal engram activates fear memory recall
Xu Liu, S. Ramirez, Petti T. Pang, C. Puryear, A. Govindarajan, K. Deisseroth, and S. Tonegawa · 2012
Earlier work this paper cites.
Episodic reinstatement in the medial temporal lobe
B. Staresina, R. Henson, N. Kriegeskorte, and Arjen Alink · 2012
Earlier work this paper cites.
Adaptive online scheduling in storm
Leonardo Aniello, R. Baldoni, and Leonardo Querzoni · 2013
Earlier work this paper cites.
Neural context reinstatement predicts memory misattribution
S. Gershman, A. Schapiro, A. Hupbach, and K. Norman · 2013
Earlier work this paper cites.
Speech recognition with deep recurrent neural networks
Alex Graves, Abdel rahman Mohamed, and Geoffrey E. Hinton · 2013
Earlier work this paper cites.
Neurobiological foundations of an attribute model of memory
R. Kesner · 2013
Earlier work this paper cites.
A buffer model of memory encoding and temporal correlations in retrieval
Melissa Lehman and Kenneth J. Malmberg · 2013
Earlier work this paper cites.
Efficient partitioning of memory systems and its importance for memory consolidation
Alex Roxin and Stefano Fusi · 2013
Earlier work this paper cites.
Attentional and non-attentional systems in the maintenance of verbal information in working memory: the executive and phonological loops
V. Camos and P. Barrouillet · 2014
Earlier work this paper cites.
Cortical reinstatement mediates the relationship between content-specific encoding activity and subsequent recollection decisions
Alan M Gordon, Jesse Rissman, Roozbeh Kiani, and Anthony D Wagner · 2014
Earlier work this paper cites.
Replay of very early encoding representations during recollection
A. Jafarpour, L. Fuentemilla, A. Horner, W. Penny, and E. Duzel · 2014
Earlier work this paper cites.
A retrieved context account of spacing and repetition effects in free recall
Lynn L Siegel and M. Kahana · 2014
Earlier work this paper cites.
Complex synapses as efficient memory systems
M. Benna and Stefano Fusi · 2015
Earlier work this paper cites.
Dual recollection in episodic memory
C. Brainerd, C. F. Gomes, and K. Nakamura · 2015
Earlier work this paper cites.
Unified pre- and postsynaptic long-term plasticity enables reliable and flexible learning
R. P. Costa, R. Froemke, P. J. Sjöström, and Mark C. W. van Rossum · 2015
Earlier work this paper cites.
Formation and maintenance of robust long-term information storage in the presence of synaptic turnover
M. Fauth, F. Wörgötter, and Christian Tetzlaff · 2015
Earlier work this paper cites.
Memory and information processing in neuromorphic systems
G. Indiveri and Shih-Chii Liu · 2015
Earlier work this paper cites.
A survey of agent platforms
K. Kravari and Nick Bassiliades · 2015
Earlier work this paper cites.
Executive resources and item-context binding: Exploring the influence of concurrent inhibition, updating, and shifting tasks on context memory
M. Nieznański, Michał Obidziński, Emilia Zyskowska, and Daria Niedziałkowska · 2015
Earlier work this paper cites.
Temporal-pattern similarity analysis reveals the beneficial and detrimental effects of context reinstatement on human memory
T. Staudigl, C. Vollmar, S. Noachtar, and S. Hanslmayr · 2015
Earlier work this paper cites.
Neural modeling of sequential inferences and learning over episodic memory
Budhitama Subagdja and A. Tan · 2015
Earlier work this paper cites.
A streaming real-time web observatory architecture for monitoring the health of social machines
Ramine Tinati, Xin Wang, Ian C. Brown, T. Tiropanis, and W. Hall · 2015
Earlier work this paper cites.
How does intentionality of encoding affect memory for episodic information?
Michael Craig, Karla Butterworth, Jonna Nilsson, Colin J Hamilton, P. Gallagher, and T. Smulders · 2016
Earlier work this paper cites.
Distributed constraint optimization problems and applications: A survey
Ferdinando Fioretto, Enrico Pontelli, and W. Yeoh · 2016
Earlier work this paper cites.
A cognitive model based on neuromodulated plasticity
Jing Huang, X. Ruan, Naigong Yu, Qingwu Fan, Jiaming Li, and Jianxian Cai · 2016
Earlier work this paper cites.
Overcoming catastrophic forgetting in neural networks
J. Kirkpatrick, Razvan Pascanu, Neil C. Rabinowitz, J. Veness, Guillaume Desjardins, Andrei A. Rusu, Kieran Milan, John Quan, Tiago Ramalho, A. Grabska-Barwinska, D. Hassabis, C. Clopath, D. Kumaran, and R. Hadsell · 2016
Earlier work this paper cites.
Assessing the ability of lstms to learn syntax-sensitive dependencies
Tal Linzen, Emmanuel Dupoux, and Yoav Goldberg · 2016
Earlier work this paper cites.
Autonomous agents modelling other agents: A comprehensive survey and open problems
Stefano V. Albrecht and P. Stone · 2017
Earlier work this paper cites.
Multi-agent systems to help managing air traffic structure
R. Breil, D. Delahaye, Laurent Lapasset, and E. Feron · 2017
Earlier work this paper cites.
Efficient attention using a fixed-size memory representation
D. Britz, M. Guan, and Minh-Thang Luong · 2017
Earlier work this paper cites.
Tree memory networks for modelling long-term temporal dependencies
Tharindu Fernando, Simon Denman, A. Mcfadyen, S. Sridharan, and C. Fookes · 2017
Earlier work this paper cites.
Model-agnostic meta-learning for fast adaptation of deep networks
Chelsea Finn, P. Abbeel, and S. Levine · 2017
Earlier work this paper cites.
Refresh my memory: Episodic memory reinstatements intrude on working memory maintenance
A. N. Hoskin, A. Bornstein, K. Norman, and J. Cohen · 2017
Earlier work this paper cites.
Billion-scale similarity search with gpus
Jeff Johnson, Matthijs Douze, and H. Jégou · 2017
Earlier work this paper cites.
The narrativeqa reading comprehension challenge
Tomás Kociský, Jonathan Schwarz, Phil Blunsom, Chris Dyer, Karl Moritz Hermann, Gábor Melis, and Edward Grefenstette · 2017
Earlier work this paper cites.
David Lillis · 2017
Earlier work this paper cites.
On the state of the art of evaluation in neural language models
Gábor Melis, Chris Dyer, and Phil Blunsom · 2017
Earlier work this paper cites.
Concepts ontology algebras and role descriptions
C. Nourani and P. Eklund · 2017
Earlier work this paper cites.
Natural tts synthesis by conditioning wavenet on mel spectrogram predictions
Jonathan Shen, Ruoming Pang, Ron J. Weiss, M. Schuster, N. Jaitly, Zongheng Yang, Z. Chen, Yu Zhang, Yuxuan Wang, R. Skerry-Ryan, R. Saurous, Yannis Agiomyrgiannakis, and Yonghui Wu · 2017
Earlier work this paper cites.
Multi-task neural network for non-discrete attribute prediction in knowledge graphs
Yi Tay, Anh Tuan Luu, Minh C. Phan, and S. Hui · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam M. Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
Tallyqa: Answering complex counting questions
Manoj Acharya, Kushal Kafle, and Christopher Kanan · 2018
Earlier work this paper cites.
Weakly-supervised 3d hand pose estimation from monocular rgb images
Yujun Cai, Liuhao Ge, Jianfei Cai, and Junsong Yuan · 2018
Earlier work this paper cites.
Question answering by reasoning across documents with graph convolutional networks
Nicola De Cao, Wilker Aziz, and Ivan Titov · 2018
Earlier work this paper cites.
The secret sharer: Evaluating and testing unintended memorization in neural networks
Nicholas Carlini, Chang Liu, Ú. Erlingsson, Jernej Kos, and D. Song · 2018
Earlier work this paper cites.
Multi-agent systems: A survey
A. Dorri, S. Kanhere, and R. Jurdak · 2018
Earlier work this paper cites.
Bilevel programming for hyperparameter optimization and meta-learning
Luca Franceschi, P. Frasconi, Saverio Salzo, Riccardo Grazzi, and M. Pontil · 2018
Earlier work this paper cites.
Reactivated spatial context guides episodic recall
Nora A. Herweg, A. Sharan, M. Sperling, A. Brandt, A. Schulze-Bonhage, and M. Kahana · 2018
Earlier work this paper cites.
A survey and analysis of cooperative multi-agent robot systems: Challenges and directions
Z. Ismail and N. Sariff · 2018
Earlier work this paper cites.
Neural tangent kernel: Convergence and generalization in neural networks
Arthur Jacot, Franck Gabriel, and Clément Hongler · 2018
Earlier work this paper cites.
T. S. Jayram, Younes Bouhadjar, Ryan L. McAvoy, Tomasz Kornuta, Alexis Asseman, K. Rocki, and A. Ozcan · 2018
Earlier work this paper cites.
Standard model of mind: Episodic memory
T. Kelley, R. Thomson, and Jonathan Milton · 2018
Earlier work this paper cites.
Reinforcement learning on web interfaces using workflow-guided exploration
E. Liu, Kelvin Guu, Panupong Pasupat, Tianlin Shi, and Percy Liang · 2018
Earlier work this paper cites.
Virtualhome: Simulating household activities via programs
Xavier Puig, K. Ra, Marko Boben, Jiaman Li, Tingwu Wang, S. Fidler, and A. Torralba · 2018
Earlier work this paper cites.
Machine theory of mind
Neil C. Rabinowitz, Frank Perbet, H. F. Song, Chiyuan Zhang, S. Eslami, and M. Botvinick · 2018
Earlier work this paper cites.
Learning quick fixes from code repositories
Reudismam Rolim, Gustavo Soares, Rohit Gheyi, and Loris D’antoni · 2018
Earlier work this paper cites.
Federico Rossi, Saptarshi Bandyopadhyay, Michael T. Wolf, and M. Pavone · 2018
Earlier work this paper cites.
Meta-transfer learning for few-shot learning
Qianru Sun, Yaoyao Liu, Tat-Seng Chua, and B. Schiele · 2018
Earlier work this paper cites.
Improving natural language inference using external knowledge in the science questions domain
Xiaoyang Wang, Pavan Kapanipathi, Ryan Musa, Mo Yu, Kartik Talamadupula, I. Abdelaziz, Maria Chang, Achille Fokoue, B. Makni, Nicholas Mattei, and M. Witbrock · 2018
Earlier work this paper cites.
Graph convolutional networks for text classification
Liang Yao, Chengsheng Mao, and Yuan Luo · 2018
Earlier work this paper cites.
Exploiting spatial-temporal relationships for 3d pose estimation via graph convolutional networks
Yujun Cai, Liuhao Ge, Jun Liu, Jianfei Cai, Tat-Jen Cham, Junsong Yuan, and Nadia Magnenat Thalmann · 2019
Earlier work this paper cites.
The nature of the traces and the dynamics of memory
Brouillet Denis and Versace Rémy · 2019
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Earlier work this paper cites.
Generalization of reinforcement learners with working and episodic memory
Meire Fortunato, Melissa Tan, Ryan Faulkner, S. Hansen, Adrià Puigdomènech Badia, Gavin Buttimore, Charlie Deck, Joel Z. Leibo, and C. Blundell · 2019
Earlier work this paper cites.
Context definition and query language: Conceptual specification, implementation, and evaluation
A. Hassani, A. Medvedev, P. D. Haghighi, Sea Ling, A. Zaslavsky, and P. Jayaraman · 2019
Earlier work this paper cites.
Examining the episodic context account: does retrieval practice enhance memory for context?
M. Hong, Sean M. Polyn, and Lisa K. Fazio · 2019
Earlier work this paper cites.
Parameter-efficient transfer learning for nlp
N. Houlsby, A. Giurgiu, Stanislaw Jastrzebski, Bruna Morrone, Quentin de Laroussilhe, Andrea Gesmundo, Mona Attariyan, and S. Gelly · 2019
Earlier work this paper cites.
Generalization through memorization: Nearest neighbor language models
Urvashi Khandelwal, Omer Levy, Dan Jurafsky, Luke Zettlemoyer, and M. Lewis · 2019
Earlier work this paper cites.
Flexible working memory through selective gating and attentional tagging
W. Kruijne, S. Bohté, P. Roelfsema, and C. Olivers · 2019
Earlier work this paper cites.
A gated self-attention memory network for answer selection
T. Lai, Quan Hung Tran, Trung Bui, and D. Kihara · 2019
Earlier work this paper cites.
Kagnet: Knowledge-aware graph networks for commonsense reasoning
Bill Yuchen Lin, Xinyue Chen, Jamin Chen, and Xiang Ren · 2019
Earlier work this paper cites.
K-bert: Enabling language representation with knowledge graph
Weijie Liu, Peng Zhou, Zhe Zhao, Zhiruo Wang, Qi Ju, Haotang Deng, and Ping Wang · 2019
Earlier work this paper cites.
Special issue “multi-agent systems”: Editorial
S. Mariani and Andrea Omicini · 2019
Earlier work this paper cites.
Right for the wrong reasons: Diagnosing syntactic heuristics in natural language inference
R. Thomas McCoy, Ellie Pavlick, and Tal Linzen · 2019
Earlier work this paper cites.
Self-supervised learning of pretext-invariant representations
Ishan Misra and L. Maaten · 2019
Earlier work this paper cites.
Modeling the relationship between identifier name and behavior
Christian D. Newman, Anthony S Peruma, and Reem S. Alsuhaibani · 2019
Earlier work this paper cites.
Self-supervised contextual data augmentation for natural language processing
Dongju Park and Chang Wook Ahn · 2019
Earlier work this paper cites.
Improving coordination in small-scale multi-agent deep reinforcement learning through memory-driven communication
E. Pesce and G. Montana · 2019
Earlier work this paper cites.
Rethinking table recognition using graph neural networks
S. Qasim, Hassan Mahmood, and F. Shafait · 2019
Earlier work this paper cites.
A neuromimetic approach to the serial acquisition, long-term storage, and selective utilization of overlapping memory engrams
Victor Quintanar-Zilinskas · 2019
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam M. Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu · 2019
Earlier work this paper cites.
Are red roses red? evaluating consistency of question-answering models
Marco Tulio Ribeiro, Carlos Guestrin, and Sameer Singh · 2019
Earlier work this paper cites.
Alfred: A benchmark for interpreting grounded instructions for everyday tasks
Mohit Shridhar, Jesse Thomason, Daniel Gordon, Yonatan Bisk, Winson Han, Roozbeh Mottaghi, Luke Zettlemoyer, and D. Fox · 2019
Earlier work this paper cites.
Ernie 2.0: A continual pre-training framework for language understanding
Yu Sun, Shuohuan Wang, Yukun Li, Shikun Feng, Hao Tian, Hua Wu, and Haifeng Wang · 2019
Earlier work this paper cites.
Self-organizing neural networks for universal learning and multimodal memory encoding
A. Tan, Budhitama Subagdja, Di Wang, and Lei Meng · 2019
Earlier work this paper cites.
Graph transformer networks
Seongjun Yun, Minbyul Jeong, Raehyun Kim, Jaewoo Kang, and Hyunwoo J. Kim · 2019
Earlier work this paper cites.
Large scale knowledge graph based synthetic corpus generation for knowledge-enhanced language model pre-training
Oshin Agarwal, Heming Ge, Siamak Shakeri, and Rami Al-Rfou · 2020
Earlier work this paper cites.
In-time explainability in multi-agent systems: Challenges, opportunities, and roadmap
Francesco Alzetta, P. Giorgini, A. Najjar, M. Schumacher, and Davide Calvaresi · 2020
Earlier work this paper cites.
Palm: Pre-training an autoencoding&autoregressive language model for context-conditioned generation
Bin Bi, Chenliang Li, Chen Wu, Ming Yan, and Wei Wang · 2020
Earlier work this paper cites.
Language models are few-shot learners
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, J. Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, T. Henighan, R. Child, A. Ramesh, Daniel M. Ziegler, Jeff Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Ma teusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, I. Sutskever, and Dario Amodei · 2020
Earlier work this paper cites.
3d hand pose estimation using synthetic data and weakly labeled rgb images
Yujun Cai, Liuhao Ge, Jianfei Cai, Nadia Magnenat Thalmann, and Junsong Yuan · 2020
Earlier work this paper cites.
Learning progressive joint propagation for human motion prediction
Yujun Cai, Lin Huang, Yiwei Wang, Tat-Jen Cham, Jianfei Cai, Junsong Yuan, Jun Liu, Xu Yang, Yiheng Zhu, Xiaohui Shen, et al · 2020
Earlier work this paper cites.
Extracting training data from large language models
Nicholas Carlini, Florian Tramèr, Eric Wallace, Matthew Jagielski, Ariel Herbert-Voss, Katherine Lee, Adam Roberts, Tom B. Brown, D. Song, Ú. Erlingsson, Alina Oprea, and Colin Raffel · 2020
Earlier work this paper cites.
A simple framework for contrastive learning of visual representations
Ting Chen, Simon Kornblith, Mohammad Norouzi, and Geoffrey E. Hinton · 2020
Earlier work this paper cites.
Rethinking attention with performers
K. Choromanski, Valerii Likhosherstov, David Dohan, Xingyou Song, Andreea Gane, Tamás Sarlós, Peter Hawkins, Jared Davis, Afroz Mohiuddin, Lukasz Kaiser, David Belanger, Lucy J. Colwell, and Adrian Weller · 2020
Earlier work this paper cites.
Self-training improves pre-training for natural language understanding
Jingfei Du, Edouard Grave, Beliz Gunel, Vishrav Chaudhary, Onur Çelebi, Michael Auli, Ves Stoyanov, and Alexis Conneau · 2020
Earlier work this paper cites.
Scalable multi-hop relational reasoning for knowledge-aware question answering
Yanlin Feng, Xinyue Chen, Bill Yuchen Lin, Peifeng Wang, Jun Yan, and Xiang Ren · 2020
Earlier work this paper cites.
Transformer feed-forward layers are key-value memories
Mor Geva, R. Schuster, Jonathan Berant, and Omer Levy · 2020
Earlier work this paper cites.
Realm: Retrieval-augmented language model pre-training
Kelvin Guu, Kenton Lee, Zora Tung, Panupong Pasupat, and Ming-Wei Chang · 2020
Earlier work this paper cites.
Multi-agent systems and complex networks: Review and applications in systems engineering
M. Herrera, Marco Pérez-Hernández, A. Kumar Parlikad, and J. Izquierdo · 2020
Earlier work this paper cites.
Knowledge graphs
Aidan Hogan, E. Blomqvist, Michael Cochez, C. d’Amato, Gerard de Melo, C. Gutierrez, J. E. L. Gayo, S. Kirrane, S. Neumaier, A. Polleres, Roberto Navigli, A. Ngomo, S. M. Rashid, Anisa Rula, Lukas Schmelzeisen, Juan Sequeda, Steffen Staab, and Antoine Zimmermann · 2020
Earlier work this paper cites.
Meta-learning in neural networks: A survey
Timothy M. Hospedales, Antreas Antoniou, P. Micaelli, and A. Storkey · 2020
Earlier work this paper cites.
Voronoi-based multi-robot autonomous exploration in unknown environments via deep reinforcement learning
Junyan Hu, Hanlin Niu, J. Carrasco, B. Lennox, and F. Arvin · 2020
Earlier work this paper cites.
Multi-agent systems: A review study
H. Jaleel, Jane J. Stephan, and Sinan Naji · 2020
Earlier work this paper cites.
A survey on knowledge graphs: Representation, acquisition, and applications
Shaoxiong Ji, Shirui Pan, E. Cambria, P. Marttinen, and Philip S. Yu · 2020
Earlier work this paper cites.
Dense passage retrieval for open-domain question answering
Vladimir Karpukhin, Barlas Oğuz, Sewon Min, Patrick Lewis, Ledell Yu Wu, Sergey Edunov, Danqi Chen, and Wen tau Yih · 2020
Earlier work this paper cites.
Transformers are rnns: Fast autoregressive transformers with linear attention
Angelos Katharopoulos, Apoorv Vyas, Nikolaos Pappas, and Franccois Fleuret · 2020
Earlier work this paper cites.
Supervised contrastive learning
Prannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna, Yonglong Tian, Phillip Isola, Aaron Maschinot, Ce Liu, and Dilip Krishnan · 2020
Earlier work this paper cites.
Reformer: The efficient transformer
Nikita Kitaev, Lukasz Kaiser, and Anselm Levskaya · 2020
Earlier work this paper cites.
Metareasoning structures, problems, and modes for multiagent systems: A survey
Samuel T. Langlois, Oghenetekevwe Akoroda, Estefany Carrillo, J. Herrmann, S. Azarm, Huan Xu, and Michael W. Otte · 2020
Earlier work this paper cites.
Self-attentive associative memory
Hung Le, T. Tran, and S. Venkatesh · 2020
Earlier work this paper cites.
Retrieval-augmented generation for knowledge-intensive nlp tasks
Patrick Lewis, Ethan Perez, Aleksandara Piktus, F. Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Kuttler, M. Lewis, Wen tau Yih, Tim Rocktäschel, Sebastian Riedel, and Douwe Kiela · 2020
Earlier work this paper cites.
Mapping natural language instructions to mobile ui action sequences
Yang Li, Jiacong He, Xiaoxia Zhou, Yuan Zhang, and Jason Baldridge · 2020
Earlier work this paper cites.
Tensor networks for language modeling
Jacob Miller, Guillaume Rabusseau, and John Terilla · 2020
Earlier work this paper cites.
A review of platforms for the development of agent systems
Constantin-Valentin Pal, F. Leon, M. Paprzycki, and M. Ganzha · 2020
Earlier work this paper cites.
Learning from context or names? an empirical study on neural relation extraction
Hao Peng, Tianyu Gao, Xu Han, Yankai Lin, Peng Li, Zhiyuan Liu, Maosong Sun, and Jie Zhou · 2020
Earlier work this paper cites.
How context affects language models’ factual predictions
F. Petroni, Patrick Lewis, Aleksandra Piktus, Tim Rocktäschel, Yuxiang Wu, Alexander H. Miller, and Sebastian Riedel · 2020
Earlier work this paper cites.
Answering open-domain questions of varying reasoning steps from text
Peng Qi, Haejun Lee, OghenetegiriTGSido, and Christopher D. Manning · 2020
Earlier work this paper cites.
Umberto-mtsa @ accompl-it: Improving complexity and acceptability prediction with multi-task learning on self-supervised annotations (short paper)
Gabriele Sarti · 2020
Earlier work this paper cites.
Learning contextual representations for semantic parsing with generation-augmented pre-training
Peng Shi, Patrick Ng, Zhiguo Wang, Henghui Zhu, Alexander Hanbo Li, Jun Wang, C. D. Santos, and Bing Xiang · 2020
Earlier work this paper cites.
Fourier features let networks learn high frequency functions in low dimensional domains
Matthew Tancik, Pratul P. Srinivasan, B. Mildenhall, Sara Fridovich-Keil, N. Raghavan, Utkarsh Singhal, R. Ramamoorthi, J. Barron, and Ren Ng · 2020
Earlier work this paper cites.
Using argumentation schemes to find motives and intentions of a rational agent
D. Walton · 2020
Earlier work this paper cites.
Spatten: Efficient sparse attention architecture with cascade token and head pruning
Hanrui Wang, Zhekai Zhang, and Song Han · 2020
Earlier work this paper cites.
Understanding graph embedding methods and their applications
Mengjia Xu · 2020
Earlier work this paper cites.
Tabert: Pretraining for joint understanding of textual and tabular data
Pengcheng Yin, Graham Neubig, Wen tau Yih, and Sebastian Riedel · 2020
Earlier work this paper cites.
Big bird: Transformers for longer sequences
M. Zaheer, Guru Guruganesh, Kumar Avinava Dubey, J. Ainslie, Chris Alberti, Santiago Ontañón, Philip Pham, Anirudh Ravula, Qifan Wang, Li Yang, and Amr Ahmed · 2020
Earlier work this paper cites.
Deepemd: Few-shot image classification with differentiable earth mover’s distance and structured classifiers
Chi Zhang, Yujun Cai, Guosheng Lin, and Chunhua Shen · 2020
Earlier work this paper cites.
On the naming of methods: A survey of professional developers
Reem S. Alsuhaibani, Christian D. Newman, M. J. Decker, Michael L. Collard, and Jonathan I. Maletic · 2021
Earlier work this paper cites.
Improving language models by retrieving from trillions of tokens
Sebastian Borgeaud, A. Mensch, Jordan Hoffmann, Trevor Cai, Eliza Rutherford, Katie Millican, George van den Driessche, Jean-Baptiste Lespiau, Bogdan Damoc, Aidan Clark, Diego de Las Casas, Aurelia Guy, Jacob Menick, Roman Ring, T. Hennigan, Saffron Huang, Lorenzo Maggiore, Chris Jones, Albin Cassirer, Andy Brock, Michela Paganini, G. Irving, O. Vinyals, Simon Osindero, K. Simonyan, Jack W. Rae, Erich Elsen, and L. Sifre · 2021
Earlier work this paper cites.
A unified 3d human motion synthesis model via conditional variational auto-encoder
Yujun Cai, Yiwei Wang, Yiheng Zhu, Tat-Jen Cham, Jianfei Cai, Junsong Yuan, Jun Liu, Chuanxia Zheng, Sijie Yan, Henghui Ding, et al · 2021
Earlier work this paper cites.
A review of agent-based programming for multi-agent systems
R. C. Cardoso and Angelo Ferrando · 2021
Earlier work this paper cites.
Meta-learning via language model in-context tuning
Yanda Chen, Ruiqi Zhong, Sheng Zha, G. Karypis, and He He · 2021
Earlier work this paper cites.
A fuzzy semantic for bdi logic
A. Cruz, André V. dos Santos, R. Santiago, and B. Bedregal · 2021
Earlier work this paper cites.
A model of semantic completion in generative episodic memory
Zahra Fayyaz, Aya Altamimi, Sen Cheng, and Laurenz Wiskott · 2021
Earlier work this paper cites.
Memory for spatio-temporal contextual details during the retrieval of naturalistic episodes
Samy Foudil, Claire Pleche, and E. Macaluso · 2021
Earlier work this paper cites.
A practical survey on faster and lighter transformers
Quentin Fournier, G. Caron, and D. Aloise · 2021
Earlier work this paper cites.
Memory capacity of neural network models, arXiv preprint arXiv:2108.07839, 2021
Stefano Fusi · 2021
Earlier work this paper cites.
Baby intuitions benchmark (bib): Discerning the goals, preferences, and actions of others
Kanishk Gandhi, Gala Stojnic, B. Lake, and M. Dillon · 2021
Earlier work this paper cites.
Making pre-trained language models better few-shot learners
Tianyu Gao, Adam Fisch, and Danqi Chen · 2021
Earlier work this paper cites.
Perceptual score: What data modalities does your model perceive?
Itai Gat, Idan Schwartz, and A. Schwing · 2021
Earlier work this paper cites.
Retroactive interference model of forgetting
Antonios Georgiou, M. Katkov, and M. Tsodyks · 2021
Earlier work this paper cites.
One thousand and one stories: a large-scale survey of software refactoring
Yaroslav Golubev, Zarina Kurbatova, E. Alomar, T. Bryksin, and Mohamed Wiem Mkaouer · 2021
Earlier work this paper cites.
Multi-agent deep reinforcement learning: a survey
Sven Gronauer and K. Diepold · 2021
Earlier work this paper cites.
Efficiently modeling long sequences with structured state spaces
Albert Gu, Karan Goel, and Christopher R’e · 2021
Earlier work this paper cites.
Elsa: Hardware-software co-design for efficient, lightweight self-attention mechanism in neural networks
Tae Jun Ham, Yejin Lee, Seong Hoon Seo, Soo-Uck Kim, Hyunji Choi, Sungjun Jung, and Jae W. Lee · 2021
Earlier work this paper cites.
A model of working memory for latent representations
Shekoofeh Hedayati, Ryan E. O’Donnell, and Brad Wyble · 2021
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
J. E. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, and Weizhu Chen · 2021
Earlier work this paper cites.
A few more examples may be worth billions of parameters
Yuval Kirstain, Patrick Lewis, Sebastian Riedel, and Omer Levy · 2021
Earlier work this paper cites.
The power of scale for parameter-efficient prompt tuning
Brian Lester, Rami Al-Rfou, and Noah Constant · 2021
Earlier work this paper cites.
Data augmentation approaches in natural language processing: A survey
Bohan Li, Yutai Hou, and Wanxiang Che · 2021
Earlier work this paper cites.
Common sense beyond english: Evaluating and improving multilingual language models for commonsense reasoning
Bill Yuchen Lin, Seyeon Lee, Xiaoyang Qiao, and Xiang Ren · 2021
Earlier work this paper cites.
Sanger: A co-design framework for enabling sparse attention using reconfigurable architecture
Liqiang Lu, Yicheng Jin, Hangrui Bi, Zizhang Luo, Peng Li, Tao Wang, and Yun Liang · 2021
Earlier work this paper cites.
Deep reinforcement learning versus evolution strategies: A comparative survey
Amjad Yousef Majid, Serge Saaybi, Tomas van Rietbergen, Vincent François-Lavet, R. V. Prasad, and Chris Verhoeven · 2021
Earlier work this paper cites.
Gnn-lm: Language modeling based on global contexts via gnn
Yuxian Meng, Shi Zong, Xiaoya Li, Xiaofei Sun, Tianwei Zhang, Fei Wu, and Jiwei Li · 2021
Earlier work this paper cites.
Context memory encoding and retrieval temporal dynamics are modulated by attention across the adult lifespan
Soroush Mirjalili, Patrick S. Powell, Jonathan Strunk, Taylor A James, and Audrey Duarte · 2021
Earlier work this paper cites.
Train short, test long: Attention with linear biases enables input length extrapolation
Ofir Press, Noah A. Smith, and M. Lewis · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, A. Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and I. Sutskever · 2021
Earlier work this paper cites.
High-resolution image synthesis with latent diffusion models
Robin Rombach, A. Blattmann, Dominik Lorenz, Patrick Esser, and B. Ommer · 2021
Earlier work this paper cites.
Federico Rossi, Saptarshi Bandyopadhyay, Michael T. Wolf, and M. Pavone · 2021
Earlier work this paper cites.
Sparsebert: Rethinking the importance analysis in self-attention
Han Shi, Jiahui Gao, Xiaozhe Ren, Hang Xu, Xiaodan Liang, Zhenguo Li, and J. Kwok · 2021
Earlier work this paper cites.
Text data augmentation for deep learning
Connor Shorten, T. Khoshgoftaar, and B. Furht · 2021
Earlier work this paper cites.
Paradigms of computational agency
S. Srinivasa and Jayati Deshmukh · 2021
Earlier work this paper cites.
Bernie Wang, Si ting Xu, K. Keutzer, Yang Gao, and Bichen Wu · 2021
Earlier work this paper cites.
Mixup for node and graph classification
Yiwei Wang, Wei Wang, Yuxuan Liang, Yujun Cai, and Bryan Hooi · 2021
Earlier work this paper cites.
Multi-agent reinforcement learning based distributed transmission in collaborative cloud-edge systems
Chunmei Xu, Shengheng Liu, Cheng Zhang, Yongming Huang, Zhaohua Lu, and Luxi Yang · 2021
Earlier work this paper cites.
Graphformers: Gnn-nested transformers for representation learning on textual graph
Junhan Yang, Zheng Liu, Shitao Xiao, Chaozhuo Li, Defu Lian, Sanjay Agrawal, Amit Singh, Guangzhong Sun, and Xing Xie · 2021
Earlier work this paper cites.
Qa-gnn: Reasoning with language models and knowledge graphs for question answering
Michihiro Yasunaga, Hongyu Ren, Antoine Bosselut, Percy Liang, and J. Leskovec · 2021
Earlier work this paper cites.
Contextual prediction errors reorganize naturalistic episodic memories in time
Fahd Yazin, Moumita Das, A. Banerjee, and Dipanjan Roy · 2021
Earlier work this paper cites.
How working memory and reinforcement learning are intertwined: A cognitive, neural, and computational perspective
Aspen H. Yoo and A. Collins · 2021
Earlier work this paper cites.
Physical safety and cyber security analysis of multi-agent systems: A survey of recent advances
Dan Zhang, G. Feng, Yang Shi, and D. Srinivasan · 2021
Earlier work this paper cites.
Calibrate before use: Improving few-shot performance of language models
Tony Zhao, Eric Wallace, Shi Feng, D. Klein, and Sameer Singh · 2021
Earlier work this paper cites.
Textgnn: Improving text encoder via graph neural network in sponsored search
Jason Zhu, Yanling Cui, Yuming Liu, Hao Sun, Xue Li, Markus Pelger, Liangjie Zhang, Tianqi Yan, Ruofei Zhang, and Huasha Zhao · 2021
Earlier work this paper cites.
Automating human evaluation of dialogue systems
S. A · 2022
Earlier work this paper cites.
Large language models are few-shot clinical information extractors
Monica Agrawal, S. Hegselmann, Hunter Lang, Yoon Kim, and D. Sontag · 2022
Earlier work this paper cites.
Flamingo: a visual language model for few-shot learning
Jean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech, Iain Barr, Yana Hasson, Karel Lenc, A. Mensch, Katie Millican, Malcolm Reynolds, Roman Ring, Eliza Rutherford, Serkan Cabi, Tengda Han, Zhitao Gong, Sina Samangooei, Marianne Monteiro, Jacob Menick, Sebastian Borgeaud, Andy Brock, Aida Nematzadeh, Sahand Sharifzadeh, Mikolaj Binkowski, Ricardo Barreira, O. Vinyals, Andrew Zisserman, and K. Simonyan · 2022
Earlier work this paper cites.
Language models as agent models
Jacob Andreas · 2022
Earlier work this paper cites.
Characterizing verbatim short-term memory in neural language models
K. Armeni, C. Honey, and Tal Linzen · 2022
Earlier work this paper cites.
Constitutional ai: Harmlessness from ai feedback, arXiv preprint arXiv:2212.08073, 2022
Yuntao Bai, Saurav Kadavath, Sandipan Kundu, Amanda Askell, Jackson Kernion, Andy Jones, Anna Chen, Anna Goldie, Azalia Mirhoseini, Cameron McKinnon, Carol Chen, Catherine Olsson, Christopher Olah, Danny Hernandez, Dawn Drain, Deep Ganguli, Dustin Li, Eli Tran-Johnson, Ethan Perez, Jamie Kerr, Jared Mueller, Jeffrey Ladish, Joshua Landau, Kamal Ndousse, Kamile Lukosuite, Liane Lovitt, Michael Sellitto, Nelson Elhage, Nicholas Schiefer, Noemi Mercado, Nova DasSarma, Robert Lasenby, Robin Larson, Sam Ringer, Scott Johnston, Shauna Kravec, Sheer El Showk, Stanislav Fort, Tamera Lanham, Timothy Telleen-Lawton, Tom Conerly, Tom Henighan, Tristan Hume, Samuel R. Bowman, Zac Hatfield-Dodds, Ben Mann, Dario Amodei, Nicholas Joseph, Sam McCandlish, Tom Brown, and Jared Kaplan · 2022
Earlier work this paper cites.
Judgment aggregation, discursive dilemma and reflective equilibrium: Neural language models as self-improving doxastic agents
Gregor Betz and Kyle Richardson · 2022
Earlier work this paper cites.
Audiolm: A language modeling approach to audio generation
Zalán Borsos, Raphaël Marinier, Damien Vincent, E. Kharitonov, O. Pietquin, Matthew Sharifi, Dominik Roblek, O. Teboul, David Grangier, M. Tagliasacchi, and Neil Zeghidour · 2022
Earlier work this paper cites.
Encoding contexts are incidentally reinstated during competitive retrieval and track the temporal dynamics of memory interference
Inês Bramão, Jiefeng Jiang, A. Wagner, and M. Johansson · 2022
Earlier work this paper cites.
Large language models can implement policy iteration
Ethan A. Brooks, Logan Walls, Richard L. Lewis, and Satinder Singh · 2022
Earlier work this paper cites.
A bio-inspired implementation of a sparse-learning spike-based hippocampus memory model
Daniel Casanueva-Morato, A. Ayuso-Martinez, J. P. Dominguez-Morales, A. Jiménez-Fernandez, and G. Jiménez-Moreno · 2022
Earlier work this paper cites.
Thinking on new system for big data technology
Xueqi CHEGN, Shenghua Liu, and Ruqing ZHANG · 2022
Earlier work this paper cites.
Program of thoughts prompting: Disentangling computation from reasoning for numerical reasoning tasks
Wenhu Chen, Xueguang Ma, Xinyi Wang, and William W. Cohen · 2022
Earlier work this paper cites.
Relationprompt: Leveraging prompts to generate synthetic data for zero-shot relation triplet extraction
Yew Ken Chia, Lidong Bing, Soujanya Poria, and Luo Si · 2022
Earlier work this paper cites.
Kai Cui, Anam Tahir, Gizem Ekinci, Ahmed Elshamanhory, Yannick Eich, Mengguang Li, and H. Koeppl · 2022
Earlier work this paper cites.
Flashattention: Fast and memory-efficient exact attention with io-awareness
Tri Dao, Daniel Y. Fu, Stefano Ermon, A. Rudra, and Christopher R’e · 2022
Earlier work this paper cites.
Minedojo: Building open-ended embodied agents with internet-scale knowledge
Linxi (Jim) Fan, Guanzhi Wang, Yunfan Jiang, Ajay Mandlekar, Yuncong Yang, Haoyi Zhu, Andrew Tang, De-An Huang, Yuke Zhu, and Anima Anandkumar · 2022
Earlier work this paper cites.
An end-to-end contrastive self-supervised learning framework for language understanding
Hongchao Fang and Pengtao Xie · 2022
Earlier work this paper cites.
Precise zero-shot dense retrieval without relevance labels
Luyu Gao, Xueguang Ma, Jimmy J. Lin, and Jamie Callan · 2022
Earlier work this paper cites.
Augmentation invariant discrete representation for generative spoken language modeling
Itai Gat, Felix Kreuk, Tu Nguyen, Ann Lee, Jade Copet, Gabriel Synnaeve, Emmanuel Dupoux, and Yossi Adi · 2022
Earlier work this paper cites.
Transformer feed-forward layers build predictions by promoting concepts in the vocabulary space
Mor Geva, Avi Caciularu, Ke Wang, and Yoav Goldberg · 2022
Earlier work this paper cites.
Amelia Glaese, Nat McAleese, Maja Trębacz, John Aslanides, Vlad Firoiu, Timo Ewalds, Maribeth Rauh, Laura Weidinger, Martin Chadwick, Phoebe Thacker, Lucy Campbell-Gillingham, Jonathan Uesato, Po-Sen Huang, Ramona Comanescu, Fan Yang, Abigail See, Sumanth Dathathri, Rory Greig, Charlie Chen, Doug Fritz, Jaume Sanchez Elias, Richard Green, Soňa Mokrá, Nicholas Fernando, Boxi Wu, Rachel Foley, Susannah Young, Iason Gabriel, William Isaac, John Mellor, Demis Hassabis, Koray Kavukcuoglu, Lisa Anne Hendricks, and Geoffrey Irving · 2022
Earlier work this paper cites.
On the parameterization and initialization of diagonal state space models
Albert Gu, Ankit Gupta, Karan Goel, and Christopher Ré · 2022
Earlier work this paper cites.
Contextual inference in learning and memory
James B. Heald, M. Lengyel, and D. Wolpert · 2022
Earlier work this paper cites.
Meta-learning the difference: Preparing large language models for efficient adaptation
Zejiang Hou, Julian Salazar, and George Polovets · 2022
Earlier work this paper cites.
A survey of knowledge enhanced pre-trained language models
Linmei Hu, Zeyi Liu, Ziwang Zhao, Lei Hou, Liqiang Nie, and Juanzi Li · 2022
Earlier work this paper cites.
Language models as zero-shot planners: Extracting actionable knowledge for embodied agents
Wenlong Huang, P. Abbeel, Deepak Pathak, and Igor Mordatch · 2022
Earlier work this paper cites.
V. Ioannidis, Xiang Song, Da Zheng, Houyu Zhang, Jun Ma, Yi Xu, Belinda Zeng, Trishul M. Chilimbi, and G. Karypis · 2022
Earlier work this paper cites.
Caigao Jiang, Siqiao Xue, James Zhang, Lingyue Liu, Zhibo Zhu, and Hongyan Hao · 2022
Earlier work this paper cites.
Heterformer: Transformer-based deep node representation learning on heterogeneous text-rich networks
Bowen Jin, Yu Zhang, Qi Zhu, and Jiawei Han · 2022
Earlier work this paper cites.
Decomposed prompting: A modular approach for solving complex tasks
Tushar Khot, H. Trivedi, Matthew Finlayson, Yao Fu, Kyle Richardson, Peter Clark, and Ashish Sabharwal · 2022
Earlier work this paper cites.
Louis Kirsch, James Harrison, Jascha Narain Sohl-Dickstein, and Luke Metz · 2022
Earlier work this paper cites.
Large language models are zero-shot reasoners
Takeshi Kojima, S. Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa · 2022
Earlier work this paper cites.
Large language models with controllable working memory, arXiv preprint arXiv:2211.05110, 2022
Daliang Li, Ankit Singh Rawat, Manzil Zaheer, Xin Wang, Michal Lukasik, Andreas Veit, Felix Yu, and Sanjiv Kumar · 2022
Earlier work this paper cites.
Memory-assisted prompt editing to improve gpt-3 after deployment
Aman Madaan, Niket Tandon, Peter Clark, and Yiming Yang · 2022
Earlier work this paper cites.
Locating and editing factual associations in gpt
Kevin Meng, David Bau, A. Andonian, and Yonatan Belinkov · 2022
Earlier work this paper cites.
Skill: Structured knowledge infusion for large language models
Fedor Moiseev, Zhe Dong, Enrique Alfonseca, and Martin Jaggi · 2022
Earlier work this paper cites.
Combining theory of mind and abduction for cooperation under imperfect information
Nieves Montes, N. Osman, and C. Sierra · 2022
Earlier work this paper cites.
Reiichiro Nakano, Jacob Hilton, Suchir Balaji, Jeff Wu, Long Ouyang, Christina Kim, Christopher Hesse, Shantanu Jain, Vineet Kosaraju, William Saunders, Xu Jiang, Karl Cobbe, Tyna Eloundou, Gretchen Krueger, Kevin Button, Matthew Knight, Benjamin Chess, and John Schulman · 2022
Earlier work this paper cites.
Modularization in belief-desire-intention agent programming and artifact-based environments
Gustavo Ortiz-Hernández, Alejandro Guerra-Hernández, J. Hübner, and W. A. Luna-Ramírez · 2022
Earlier work this paper cites.
Talm: Tool augmented language models
A Parisi, Y Zhao, and N Fiedel · 2022
Earlier work this paper cites.
Social simulacra: Creating populated prototypes for social computing systems
J. Park, Lindsay Popowski, Carrie J. Cai, M. Morris, Percy Liang, and Michael S. Bernstein · 2022
Earlier work this paper cites.
Binjie Qin, Haohao Mao, Ruipeng Zhang, Y. Zhu, Song Ding, and Xu Chen · 2022
Earlier work this paper cites.
S. Rizvi, Nazreen Pallikkavaliyaveetil, David Zhang, Zhuoyang Lyu, Nhi Nguyen, Haoran Lyu, B. Christensen, J. O. Caro, Antonio H. O. Fonseca, E. Zappala, Maryam Bagherian, Christopher Averill, C. Abdallah, Amin Karbasi, Rex Ying, M. Brbic, R. M. Dhodapkar, and David van Dijk · 2022
Earlier work this paper cites.
Leveraging large language models for multiple choice question answering
Joshua Robinson, Christopher Rytting, and D. Wingate · 2022
Earlier work this paper cites.
Sequence-to-sequence knowledge graph completion and question answering
Apoorv Saxena, Adrian Kochsiek, and Rainer Gemulla · 2022
Earlier work this paper cites.
On the effect of pretraining corpora on in-context learning by a large-scale language model
Seongjin Shin, Sang-Woo Lee, Hwijeen Ahn, Sungdong Kim, Hyoungseok Kim, Boseop Kim, Kyunghyun Cho, Gichang Lee, W. Park, Jung-Woo Ha, and Nako Sung · 2022
Earlier work this paper cites.
Llm-planner: Few-shot grounded planning for embodied agents with large language models
Chan Hee Song, Jiaman Wu, Clay Washington, Brian M. Sadler, Wei-Lun Chao, and Yu Su · 2022
Earlier work this paper cites.
Meta-gui: Towards multi-modal conversational agents on mobile gui
Liangtai Sun, Xingyu Chen, Lu Chen, Tianle Dai, Zichen Zhu, and Kai Yu · 2022
Earlier work this paper cites.
Xuemei Tang, Jun Wang, and Q. Su · 2022
Earlier work this paper cites.
Memorization without overfitting: Analyzing the training dynamics of large language models
Kushal Tirumala, Aram H. Markosyan, Luke Zettlemoyer, and Armen Aghajanyan · 2022
Earlier work this paper cites.
A model of working memory for encoding multiple items and ordered sequences exploiting the theta-gamma code
M. Ursino, Nicole Cesaretti, and G. Pirazzini · 2022
Earlier work this paper cites.
Expectation vs. experience: Evaluating the usability of code generation tools powered by large language models
Priyan Vaithilingam, Tianyi Zhang, and Elena L. Glassman · 2022
Earlier work this paper cites.
Visually-augmented language modeling
Weizhi Wang, Li Dong, Hao Cheng, Haoyu Song, Xiaodong Liu, Xifeng Yan, Jianfeng Gao, and Furu Wei · 2022
Earlier work this paper cites.
Chain of thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Ed H. Chi, F. Xia, Quoc Le, and Denny Zhou · 2022
Earlier work this paper cites.
Webshop: Towards scalable real-world web interaction with grounded language agents
Shunyu Yao, Howard Chen, John Yang, and Karthik Narasimhan · 2022
Earlier work this paper cites.
Deep bidirectional language-knowledge graph pretraining
Michihiro Yasunaga, Antoine Bosselut, Hongyu Ren, Xikun Zhang, Christopher D. Manning, Percy Liang, and J. Leskovec · 2022
Earlier work this paper cites.
Prompt engineering for zero-shot and few-shot defect detection and classification using a visual-language pretrained model
Gunwoo Yong, Kahyun Jeon, Daeyoung Gil, and Ghang Lee · 2022
Earlier work this paper cites.
Vision transformer based on knowledge distillation in tcm image classification
Ge Yuyao, Cheng Yiting, Wang Jia, Zhou Hanlin, and Chen Lizhe · 2022
Earlier work this paper cites.
Star: Bootstrapping reasoning with reasoning, arXiv preprint arXiv:2203.14465, 2022
E. Zelikman, Yuhuai Wu, and Noah D. Goodman · 2022
Earlier work this paper cites.
Subgraph retrieval enhanced model for multi-hop knowledge base question answering
Jing Zhang, Xiaokang Zhang, Jifan Yu, Jian Tang, Jie Tang, Cuiping Li, and Hong Chen · 2022
Earlier work this paper cites.
Learning on large-scale text-attributed graphs via variational inference
Jianan Zhao, Meng Qu, Chaozhuo Li, Hao Yan, Qian Liu, Rui Li, Xing Xie, and Jian Tang · 2022
Earlier work this paper cites.
Semantic-aware event link reasoning over industrial knowledge graph embedding time series data
Bin Zhou, Xingwang Shen, Yuqian Lu, Xinyu Li, B. Hua, Tianyuan Liu, and Jinsong Bao · 2022
Earlier work this paper cites.
Kitlm: Domain-specific knowledge integration into language models for question answering
Ankush Agarwal, Sakharam Gawade, A. Azad, and P. Bhattacharyya · 2023
Earlier work this paper cites.
Gqa: Training generalized multi-query transformer models from multi-head checkpoints
J. Ainslie, J. Lee-Thorp, Michiel de Jong, Yury Zemlyanskiy, Federico Lebr’on, and Sumit K. Sanghai · 2023
Earlier work this paper cites.
Position interpolation improves alibi extrapolation
Faisal Al-Khateeb, Nolan Dey, Daria Soboleva, and Joel Hestness · 2023
Earlier work this paper cites.
Dynamic context pruning for efficient and interpretable autoregressive transformers
Sotiris Anagnostidis, Dario Pavllo, Luca Biggio, Lorenzo Noci, Aurélien Lucchi, and Thomas Hofmann · 2023
Earlier work this paper cites.
Learning reward machines in cooperative multi-agent tasks
Leo Ardon, Daniel Furelos-Blanco, and A. Russo · 2023
Earlier work this paper cites.
Self-rag: Learning to retrieve, generate, and critique through self-reflection
Akari Asai, Zeqiu Wu, Yizhong Wang, Avirup Sil, and Hannaneh Hajishirzi · 2023
Earlier work this paper cites.
Bobby Azad, Reza Azad, Sania Eskandari, Afshin Bozorgpour, A. Kazerouni, I. Rekik, and D. Merhof · 2023
Earlier work this paper cites.
Transformers for tabular data representation: A survey of models and applications
Gilbert Badaro, Mohammed Saeed, and Paolo Papotti · 2023
Earlier work this paper cites.
Knowledge-augmented large language models for personalized contextual query suggestion
Jinheon Baek, N. Chandrasekaran, Silviu Cucerzan, Allen Herring, and S. Jauhar · 2023
Earlier work this paper cites.
Globaldoc: A cross-modal vision-language framework for real-world document image retrieval and classification
Souhail Bakkali, Sanket Biswas, Zuheng Ming, Mickaël Coustaty, Marccal Rusinol, O. R. Terrades, and Josep Llad’os · 2023
Earlier work this paper cites.
Towards hybrid automation by bootstrapping conversational interfaces for it operation tasks
Jayachandu Bandlamudi, K. Mukherjee, Prerna Agarwal, Sampath Dechu, Siyu Huo, Vatche Isahagian, Vinod Muthusamy, N. Purushothaman, and Renuka Sindhgatta · 2023
Earlier work this paper cites.
Ask and you shall receive (a graph drawing): Testing chatgpt’s potential to apply graph layout algorithms
Sara Di Bartolomeo, Giorgio Severi, V. Schetinger, and Cody Dunne · 2023
Earlier work this paper cites.
Unlimiformer: Long-range transformers with unlimited length input
Amanda Bertsch, Uri Alon, Graham Neubig, and Matthew R. Gormley · 2023
Earlier work this paper cites.
Graph of thoughts: Solving elaborate problems with large language models
Maciej Besta, Nils Blach, Aleš Kubíček, Robert Gerstenberger, Lukas Gianinazzi, Joanna Gajda, Tomasz Lehmann, Michal Podstawski, H. Niewiadomski, P. Nyczyk, and Torsten Hoefler · 2023
Earlier work this paper cites.
Grounding mental representations in a virtual multi-level functional framework
P. Bonzon · 2023
Earlier work this paper cites.
Augmenting large language models with chemistry tools
Andrés M Bran, Sam Cox, Oliver Schilter, Carlo Baldassari, Andrew D. White, and P. Schwaller · 2023
Earlier work this paper cites.
Retentive or forgetful? diving into the knowledge memorizing mechanism of language models
Boxi Cao, Qiaoyu Tang, Hongyu Lin, Xianpei Han, Jiawei Chen, Tianshu Wang, and Le Sun · 2023
Earlier work this paper cites.
Bio-inspired computational memory model of the hippocampus: an approach to a neuromorphic spike-based content-addressable memory
Daniel Casanueva-Morato, A. Ayuso-Martinez, J. P. Dominguez-Morales, A. Jiménez-Fernandez, and G. Jiménez-Moreno · 2023
Earlier work this paper cites.
clembench: Using game play to evaluate chat-optimized language models as conversational agents
Kranti Chalamalasetti, Jana Gotze, Sherzod Hakimov, Brielen Madureira, Philipp Sadler, and David Schlangen · 2023
Earlier work this paper cites.
A survey on evaluation of large language models
Yu-Chu Chang, Xu Wang, Jindong Wang, Yuan Wu, Kaijie Zhu, Hao Chen, Linyi Yang, Xiaoyuan Yi, Cunxiang Wang, Yidong Wang, Weirong Ye, Yue Zhang, Yi Chang, Philip S. Yu, Qian Yang, and Xingxu Xie · 2023
Earlier work this paper cites.
Vlp: A survey on vision-language pre-training
Fei-Long Chen, Du-Zhen Zhang, Ming-Lun Han, Xiu-Yi Chen, Jing Shi, Shuang Xu, and Bo Xu · 2023
Earlier work this paper cites.
Large knowledge model: Perspectives and challenges
Huajun Chen · 2023
Earlier work this paper cites.
Lift yourself up: Retrieval-augmented text generation with self memory
Xin Cheng, Di Luo, Xiuying Chen, Lemao Liu, Dongyan Zhao, and Rui Yan · 2023
Earlier work this paper cites.
Data-centric financial large language models, arXiv preprint arXiv:2310.17784, 2023
Zhixuan Chu, Huaiyu Guo, Xinyuan Zhou, Yijia Wang, Fei Yu, Hong Chen, Wanqing Xu, Xin Lu, Qing Cui, Longfei Li, Junqing Zhou, and Sheng Li · 2023
Earlier work this paper cites.
Meta-in-context learning in large language models
Julian Coda-Forno, Marcel Binz, Zeynep Akata, M. Botvinick, Jane X. Wang, and Eric Schulz · 2023
Earlier work this paper cites.
Flashattention-2: Faster attention with better parallelism and work partitioning
Tri Dao · 2023
Earlier work this paper cites.
On meta-prompting, arXiv preprint arXiv:2312.06562, 2023
Adrian de Wynter, Xun Wang, Qilong Gu, and Si-Qing Chen · 2023
Earlier work this paper cites.
Mind2web: Towards a generalist agent for the web
Xiang Deng, Yu Gu, Boyuan Zheng, Shijie Chen, Samuel Stevens, Boshi Wang, Huan Sun, and Yu Su · 2023
Earlier work this paper cites.
Mohammad Mahdi Derakhshani, Ivona Najdenkoska, Cees G. M. Snoek, M. Worring, and Yuki Asano · 2023
Earlier work this paper cites.
Dhruv Dhamani and Mary Lou Maher · 2023
Earlier work this paper cites.
Longnet: Scaling transformers to 1, 000, 000, 000 tokens
Jiayu Ding, Shuming Ma, Li Dong, Xingxing Zhang, Shaohan Huang, Wenhui Wang, and Furu Wei · 2023
Earlier work this paper cites.
Revisit input perturbation problems for llms: A unified robustness evaluation framework for noisy slot filling task
Guanting Dong, Jinxu Zhao, Tingfeng Hui, Daichi Guo, Wenlong Wan, Boqi Feng, Yueyan Qiu, Zhuoma Gongque, Keqing He, Zechen Wang, and Weiran Xu · 2023
Earlier work this paper cites.
Talk like a graph: Encoding graphs for large language models
Bahare Fatemi, Jonathan J. Halcrow, and Bryan Perozzi · 2023
Earlier work this paper cites.
Constant memory attention block, arXiv preprint arXiv:2306.12599, 2023
Leo Feng, Frederick Tung, Hossein Hajimirsadeghi, Y. Bengio, and M. O. Ahmed · 2023
Earlier work this paper cites.
Promptbreeder: Self-referential self-improvement via prompt evolution
Chrisantha Fernando, Dylan Banarse, H. Michalewski, Simon Osindero, and Tim Rocktäschel · 2023
Earlier work this paper cites.
Context-aware meta-learning
Christopher Fifty, Dennis Duan, Ronald G. Junkins, Ehsan Amid, Jurij Leskovec, Christopher R’e, and Sebastian Thrun · 2023
Cited alongside, same era.
Mathematical modeling of human memory
Paolo Finotelli and Francis Eustache · 2023
Cited alongside, same era.
Sgcn: a multi-order neighborhood feature fusion landform classification method based on superpixel and graph convolutional network
Honghao Fu, Yilang Shen, Yuxuan Liu, Jingzhong Li, and Xiang Zhang · 2023
Cited alongside, same era.
Large language models empowered agent-based modeling and simulation: A survey and perspectives
Chen Gao, Xiaochong Lan, Nian Li, Yuan Yuan, Jingtao Ding, Zhilun Zhou, Fengli Xu, and Yong Li · 2023
Cited alongside, same era.
In-context autoencoder for context compression in a large language model
Tao Ge, Jing Hu, Xun Wang, Si-Qing Chen, and Furu Wei · 2023
Cited alongside, same era.
Raptor: Recursive abstractive processing for tree-organized retrieval
Parth Sarthi, Salman Abdullah, Aditi Tuli, Shubh Khanna, Anna Goldie, and Christopher D. Manning · 2024
Later among the works it cites.
Wenbo Shang and Xin Huang · 2024
Later among the works it cites.
On linearizing structured data in encoder-decoder language models: Insights from text-to-sql
Yutong Shao and N. Nakashole · 2024
Later among the works it cites.
Scribeagent: Towards specialized web agents using production-scale workflow data
Junhong Shen, Atishay Jain, Zedian Xiao, Ishan Amlekar, Mouad Hadji, Aaron Podolny, and Ameet Talwalkar · 2024
Later among the works it cites.
Llm with tools: A survey, arXiv preprint arXiv:2409.18807, 2024
Zhuocheng Shen · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Hejia Geng, Boxun Xu, and Peng Li · 2023
Cited alongside, same era.
Critic: Large language models can self-correct with tool-interactive critiquing
Z Gou, Z Shao, Y Gong, Y Shen, and Y Yang… · 2023
Cited alongside, same era.
Jiayan Guo, Lun Du, and Hengyu Liu · 2023
Cited alongside, same era.
A real-world webagent with planning, long context understanding, and program synthesis
I Gur, H Furuta, A Huang, M Safdari, and Y Matsuo… · 2023
Cited alongside, same era.
Toolkengpt: Augmenting frozen language models with massive tools via tool embeddings
S Hao, T Liu, Z Wang, and Z Hu · 2023
Cited alongside, same era.
Prompt engineering in medical education
Thomas F. Heston and Charya Khun · 2023
Cited alongside, same era.
Comparative analysis of gpt-4 and human graders in evaluating human tutors giving praise to students
Dollaya Hirunyasiri, Danielle R. Thomas, Jionghao Lin, K. Koedinger, and Vincent Aleven · 2023
Cited alongside, same era.
Later among the works it cites.
Tool learning in the wild: Empowering language models as automatic tool agents
Zhengliang Shi, Shen Gao, Xiuyi Chen, Yue Feng, Lingyong Yan, Haibo Shi, Dawei Yin, Zhumin Chen, Suzan Verberne, and Zhaochun Ren · 2024
Later among the works it cites.
Jay Shim, Grant Kruttschnitt, Alyssa Ma, Daniel Kim, Benjamin Chek, Athul Anand, Kevin Zhu, and Sean O’Brien · 2024
Later among the works it cites.
Retrieval-augmented test generation: How far are we?, arXiv preprint arXiv:2409.12682, 2024
Jiho Shin, Reem Aleithan, Hadi Hemmati, and Song Wang · 2024
Later among the works it cites.
An empirical analysis on spatial reasoning capabilities of large multimodal models
Fatemeh Shiri, Xiao-Yu Guo, Mona Far, Xin Yu, Reza Haf, and Yuan-Fang Li · 2024
Later among the works it cites.
Toward a dynamic future with adaptable computing and network convergence (acnc)
Masoud Shokrnezhad, Hao Yu, T. Taleb, Renwei Li, Kyunghan Lee, Jaeseung Song, and Cedric Westphal · 2024
Later among the works it cites.
Aditi Singh, Abul Ehtesham, Gaurav Kumar Gupta, Nikhil Kumar Chatta, Saket Kumar, and T. T. Khoei · 2024
Later among the works it cites.
Leveraging large language models for optimized item categorization using unspsc taxonomy
Anmolika Singh and Yuhang Diao · 2024
Later among the works it cites.
Is synthetic data all we need? benchmarking the robustness of models trained with synthetic images
Krishnakant Singh, Thanush Navaratnam, Jannik Holmer, Simone Schaub-Meyer, and Stefan Roth · 2024
Later among the works it cites.
Maml-en-llm: Model agnostic meta-training of llms for improved in-context learning
Sanchit Sinha, Yuguang Yue, Victor Soto, Mayank Kulkarni, Jianhua Lu, and Aidong Zhang · 2024
Later among the works it cites.
Step: Stacked llm policies for web actions, arXiv preprint arXiv:2310.03720, 2024
Paloma Sodhi, S. R. K. Branavan, Yoav Artzi, and Ryan McDonald · 2024
Later among the works it cites.
Hierarchical context merging: Better long context understanding for pre-trained llms
Woomin Song, Seunghyuk Oh, Sangwoo Mo, Jaehyung Kim, Sukmin Yun, Jung-Woo Ha, and Jinwoo Shin · 2024
Later among the works it cites.
Chain of thoughtlessness? an analysis of cot in planning
Kaya Stechly, Karthik Valmeekam, and Subbarao Kambhampati · 2024
Later among the works it cites.
Olly Styles, Sam Miller, Patricio Cerda-Mardini, T. Guha, Victor Sanchez, and Bertie Vidgen · 2024
Later among the works it cites.
Guangxin Su, Yifan Zhu, Wenjie Zhang, Hanchen Wang, and Ying Zhang · 2024
Later among the works it cites.
Chuanneng Sun, Songjun Huang, and D. Pompili · 2024
Later among the works it cites.
Mcp-solver: Integrating language models with constraint programming systems
Stefan Szeider · 2024
Later among the works it cites.
Online adaptation of language models with a memory of amortized contexts
Jihoon Tack, Jaehyung Kim, Eric Mitchell, Jinwoo Shin, Yee Whye Teh, and Jonathan Richard Schwarz · 2024
Later among the works it cites.
Lloco: Learning long contexts offline
Sijun Tan, Xiuyu Li, Shishir G. Patil, Ziyang Wu, Tianjun Zhang, Kurt Keutzer, Joseph E. Gonzalez, and Raluca A. Popa · 2024
Later among the works it cites.
Graphgpt: Graph instruction tuning for large language models
Jiabin Tang, Yuhao Yang, Wei Wei, Lei Shi, Lixin Su, Suqi Cheng, Dawei Yin, and Chao Huang · 2024
Later among the works it cites.
Denis Tarasov and Kumar Shridhar · 2024
Later among the works it cites.
Junfeng Tian, Da Zheng, Yang Cheng, Rui Wang, Colin Zhang, and Debing Zhang · 2024
Later among the works it cites.
Evaluating the efficacy of prompt-engineered large multimodal models versus fine-tuned vision transformers in image-based security applications
Fouad Trad and Ali Chehab · 2024
Later among the works it cites.
Appworld: A controllable world of apps and people for benchmarking interactive coding agents
H. Trivedi, Tushar Khot, Mareike Hartmann, R. Manku, Vinty Dong, Edward Li, Shashank Gupta, Ashish Sabharwal, and Niranjan Balasubramanian · 2024
Later among the works it cites.
Text-centric alignment for multi-modality learning, arXiv preprint arXiv:2402.08086v2, 2024
Yun-Da Tsai, Ting-Yu Yen, Pei-Fu Guo, Zhe-Yan Li, and Shou-De Lin · 2024
Later among the works it cites.
Eduard Tulchinskii, Laida Kushnareva, Kristian Kuznetsov, Anastasia Voznyuk, Andrei Andriiainen, Irina Piontkovskaya, Evgeny Burnaev, and Serguei Barannikov · 2024
Later among the works it cites.
Phuc Phan Van, Dat Nguyen Minh, An Dinh Ngoc, and Huy-Phan Thanh · 2024
Later among the works it cites.
Gaurav Verma, Rachneet Kaur, Nishan Srishankar, Zhen Zeng, T. Balch, and Manuela Veloso · 2024
Later among the works it cites.
Aliaksei Vertsel and Mikhail Rumiantsau · 2024
Later among the works it cites.
James Vo · 2024
Later among the works it cites.
From chat to publication management: Organizing your related work using bibsonomy & llms
Tom Völker, Jan Pfister, Tobias Koopmann, and Andreas Hotho · 2024
Later among the works it cites.
Generative ai application for building industry
Hanlong Wan, Jian Zhang, Yan Chen, Weili Xu, and Fan Feng · 2024
Later among the works it cites.
Adapting llms for efficient context processing through soft prompt compression
Cangqing Wang, Yutian Yang, Ruisi Li, Dan Sun, Ruicong Cai, Yuzhu Zhang, and Chengqian Fu · 2024
Later among the works it cites.
Application of large language models based on knowledge graphs in question-answering systems: A review
Yani Wang · 2024
Later among the works it cites.
Srsa: A cost-efficient strategy-router search agent for real-world human-machine interactions
Yaqi Wang and Haipei Xu · 2024
Later among the works it cites.
Large language models are pattern matchers: Editing semi-structured and structured documents with chatgpt
Irene Weber · 2024
Later among the works it cites.
Llm-smartaudit: Advanced smart contract vulnerability detection
Zhiyuan Wei, Jing Sun, Zijian Zhang, and Xianhao Zhang · 2024
Later among the works it cites.
Cut your losses in large-vocabulary language models
Erik Wijmans, Brody Huval, Alexander Hertzberg, V. Koltun, and Philipp Krähenbühl · 2024
Later among the works it cites.
Biao Wu, Yanda Li, Meng Fang, Zirui Song, Zhiwei Zhang, Yunchao Wei, and Ling Chen · 2024
Later among the works it cites.
Xue Wu and Kostas Tsioutsiouliklis · 2024
Later among the works it cites.
Infllm: Training-free long-context extrapolation for llms with an efficient context memory
Chaojun Xiao, Pengle Zhang, Xu Han, Guangxuan Xiao, Yankai Lin, Zhengyan Zhang, Zhiyuan Liu, Song Han, and Maosong Sun · 2024
Later among the works it cites.
Travelplanner: A benchmark for real-world planning with language agents
Jian Xie, Kai Zhang, Jiangjie Chen, Tinghui Zhu, Renze Lou, Yuandong Tian, Yanghua Xiao, and Yu Su · 2024
Later among the works it cites.
Haoyi Xiong, Zhiyuan Wang, Xuhong Li, Jiang Bian, Zeke Xie, Shahid Mumtaz, and Laura E. Barnes · 2024
Later among the works it cites.
Reducing tool hallucination via reliability alignment, arXiv preprint arXiv:2412.04141, 2024
Hongshen Xu, Su Zhu, Zihan Wang, Hang Zheng, Da Ma, Ruisheng Cao, Shuai Fan, Lu Chen, and Kai Yu · 2024
Later among the works it cites.
Xiangyuan Xue, Zeyu Lu, Di Huang, Zidong Wang, Wanli Ouyang, and Lei Bai · 2024
Later among the works it cites.
The synergistic role of deep learning and neural architecture search in advancing artificial intelligence
Xu Yan, Junliang Du, Lun Wang, Yingbin Liang, Jiacheng Hu, and Bingxing Wang · 2024
Later among the works it cites.
Memory3: Language modeling with explicit memory
Hongkang Yang, Zehao Lin, Wenjin Wang, Hao Wu, Zhiyu Li, Bo Tang, Wenqiang Wei, Jinbo Wang, Zeyun Tang, Shichao Song, Chenyang Xi, Yu Yu, Kai Chen, Feiyu Xiong, Linpeng Tang, and E. Weinan · 2024
Later among the works it cites.
Adaptive control of retrieval-augmented generation for large language models through reflective tags
Chengyuan Yao and Satoshi Fujita · 2024
Later among the works it cites.
Huaiyuan Yao, Longchao Da, Vishnu Nandam, J. Turnau, Zhiwei Liu, Linsey Pang, and Hua Wei · 2024
Later among the works it cites.
J Ye, G Li, S Gao, C Huang, Y Wu, S Li, and X Fan… · 2024
Later among the works it cites.
Mmau: A holistic benchmark of agent capabilities across diverse domains
Guoli Yin, Haoping Bai, Shuang Ma, Feng Nan, Yanchao Sun, Zhaoyang Xu, Shen Ma, Jiarui Lu, Xiang Kong, Aonan Zhang, Dian Ang Yap, Yizhe Zhang, K. Ahnert, Vik Kamath, Mathias Berglund, Dominic Walsh, Tobias Gindele, Juergen Wiest, Zhengfeng Lai, Xiaoming Wang, Jiulong Shan, Meng Cao, Ruoming Pang, and Zirui Wang · 2024
Later among the works it cites.
Compact: Compressing retrieved documents actively for question answering
Chanwoong Yoon, Taewhoo Lee, Hyeon Hwang, Minbyul Jeong, and Jaewoo Kang · 2024
Later among the works it cites.
Llm-evolve: Evaluation for llm’s evolving capability on benchmarks
Jiaxuan You, Mingjie Liu, Shrimai Prabhumoye, M. Patwary, M. Shoeybi, and Bryan Catanzaro · 2024
Later among the works it cites.
Teaching llms to refine with tools, arXiv preprint arXiv:2412.16871, 2024
Dian Yu, Yuheng Zhang, Jiahao Xu, Tian Liang, Linfeng Song, Zhaopeng Tu, Haitao Mi, and Dong Yu · 2024
Later among the works it cites.
Zeping Yu and Sophia Ananiadou · 2024
Later among the works it cites.
Self-rewarding language models
Weizhe Yuan, Richard Yuanzhe Pang, Kyunghyun Cho, Sainbayar Sukhbaatar, Jing Xu, and Jason E Weston · 2024
Later among the works it cites.
Fragrel: Exploiting fragment-level relations in the external memory of large language models
Xihang Yue, Linchao Zhu, and Yi Yang · 2024
Later among the works it cites.
Pai Zeng, Zhenyu Ning, Jieru Zhao, Weihao Cui, Mengwei Xu, Liwei Guo, XuSheng Chen, and Yizhou Shan · 2024
Later among the works it cites.
Notellm-2: Multimodal large representation models for recommendation
Chao Zhang, Haoxin Zhang, Shiwei Wu, Di Wu, Tong Xu, Xiangyu Zhao, Yan Gao, Yao Hu, and Enhong Chen · 2024
Later among the works it cites.
Mm-llms: Recent advances in multimodal large language models
Duzhen Zhang, Yahan Yu, Jiahua Dong, Chenxing Li, Dan Su, Chenhui Chu, and Dong Yu · 2024
Later among the works it cites.
Hengyu Zhang · 2024
Later among the works it cites.
Expel: Llm agents are experiential learners, arXiv preprint arXiv:2308.10144, 2024
Andrew Zhao, Daniel Huang, Quentin Xu, Matthieu Lin, Yong-Jin Liu, and Gao Huang · 2024
Later among the works it cites.
Gpt-4v(ision) is a generalist web agent, if grounded
Boyuan Zheng, Boyu Gou, Jihyung Kil, Huan Sun, and Yu Su · 2024
Later among the works it cites.
Debug like a human: A large language model debugger via verifying runtime execution step by step
Li Zhong, Zilong Wang, and Jingbo Shang · 2024
Later among the works it cites.
Large language model assisted adversarial robustness neural architecture search
Rui Zhong, Yang Cao, Jun Yu, and M. Munetomo · 2024
Later among the works it cites.
Teaching-assistant-in-the-loop: Improving knowledge distillation from imperfect teacher models in low-budget scenarios
Yuhang Zhou and Wei Ai · 2024
Later among the works it cites.
Yujia Zhou, Yan Liu, Xiaoxi Li, Jiajie Jin, Hongjin Qian, Zheng Liu, Chaozhuo Li, Zhicheng Dou, Tsung-Yi Ho, and Philip S. Yu · 2024
Later among the works it cites.
Redel: A toolkit for llm-powered recursive multi-agent systems
Andrew Zhu, Liam Dugan, and Christopher Callison-Burch · 2024
Later among the works it cites.
Alex Zhuang, Ge Zhang, Tianyu Zheng, Xinrun Du, Junjie Wang, Weiming Ren, Stephen W. Huang, Jie Fu, Xiang Yue, and Wenhu Chen · 2024
Later among the works it cites.
Efficientrag: Efficient retriever for multi-hop question answering
Ziyuan Zhuang, Zhiyang Zhang, Sitao Cheng, Fangkai Yang, Jia Liu, Shujian Huang, Qingwei Lin, Saravan Rajmohan, Dongmei Zhang, and Qi Zhang · 2024
Later among the works it cites.
Triad: A framework leveraging a multi-role llm-based agent to solve knowledge base question answering
Chang Zong, Yuchen Yan, Weiming Lu, Eliot Huang, Jian Shao, and Y. Zhuang · 2024
Later among the works it cites.
https://agent-network-protocol.com/specs/communication.html
Anp-agent communication meta-protocol specification(draft) · 2025
Closest in time.
Theorem-of-thought: A multi-agent framework for abductive, deductive, and inductive reasoning in language models
Samir Abdaljalil, Hasan Kurban, Khalid A. Qaraqe, and E. Serpedin · 2025
Closest in time.
Abdelrahman Abdallah, Bhawna Piryani, Jamshid Mozafari, Mohammed Ali, and Adam Jatowt · 2025
Closest in time.
Agentic ai: Autonomous intelligence for complex goals—a comprehensive survey
D. Acharya, Karthigeyan Kuppan, and Divya Bhaskaracharya · 2025
Closest in time.
Emre Can Acikgoz, Jeremy Greer, Akul Datta, Ze Yang, William Zeng, Oussama Elachqar, Emmanouil Koukoumidis, Dilek Hakkani-Tur, and Gokhan Tur · 2025
Closest in time.
Arash Ahmadi, S. Sharif, and Yaser Mohammadi Banadaki · 2025
Closest in time.
Buthayna AlMulla, Maram Assi, and Safwat Hassan · 2025
Closest in time.
Ultraif: Advancing instruction following from the wild
Kaikai An, Li Sheng, Ganqu Cui, Shuzheng Si, Ning Ding, Yu Cheng, and Baobao Chang · 2025
Closest in time.
Introducing the model context protocol, November 2024
Anthropic · 2025
Closest in time.
RM Aratchige and Dr. Wmks Ilmini · 2025
Closest in time.
Self iterative label refinement via robust unlabeled learning, arXiv preprint arXiv:2502.12565, 2025
Hikaru Asano, Tadashi Kozuno, and Yukino Baba · 2025
Closest in time.
Sketch-of-thought: Efficient llm reasoning with adaptive cognitive-inspired sketching
Simon A. Aytes, Jinheon Baek, and Sung Ju Hwang · 2025
Closest in time.
Azadeh Beiranvand and S. M. Vahidipour · 2025
Closest in time.
Overflow prevention enhances long-context recurrent llms
Assaf Ben-Kish, Itamar Zimerman, M. J. Mirza, James R. Glass, Leonid Karlinsky, and Raja Giryes · 2025
Closest in time.
Shelly Bensal, Umar Jamil, Christopher Bryant, M. Russak, Kiran Kamble, Dmytro Mozolevskyi, Muayad Ali, and Waseem Alshikh · 2025
Closest in time.
When should we orchestrate multiple agents?, arXiv preprint arXiv:2503.13577, 2025
Umang Bhatt, Sanyam Kapoor, Mihir Upadhyay, Ilia Sucholutsky, Francesco Quinzan, Katherine M. Collins, Adrian Weller, Andrew Gordon Wilson, and Muhammad Bilal Zafar · 2025
Closest in time.
Decoding by contrasting knowledge: Enhancing llms’ confidence on edited facts
Baolong Bi, Shenghua Liu, Lingrui Mei, Yiwei Wang, Pengliang Ji, and Xueqi Cheng · 2025
Closest in time.
Refinex: Learning to refine pre-training data at scale from expert-guided programs
Baolong Bi, Shenghua Liu, Xingzhang Ren, Dayiheng Liu, Junyang Lin, Yiwei Wang, Lingrui Mei, Junfeng Fang, Jiafeng Guo, and Xueqi Cheng · 2025
Closest in time.
Is factuality enhancement a free lunch for llms? better factuality can lead to worse context-faithfulness
Baolong Bi, Shenghua Liu, Yiwei Wang, Lingrui Mei, Junfeng Fang, Hongcheng Gao, Shiyu Ni, and Xueqi Cheng · 2025
Closest in time.
Deep research bench: Evaluating ai web research agents, arXiv preprint arXiv:2506.06287, 2025
FutureSearch Nikos I. Bosse, Jon Evans, Robert G. Gambee, Daniel Hnyk, Peter Muhlbacher, Lawrence Phillips, Dan Schwarz, and Jack Wildman · 2025
Closest in time.
Agentic ai and multiagentic: Are we reinventing the wheel?, arXiv preprint arXiv:2506.01463, 2025
Vicent Botti · 2025
Closest in time.
The effectiveness of large language models in transforming unstructured text to standardized formats
William Brach, Kristián Kostál, and Michal Ries · 2025
Closest in time.
Lorenz Brehme, Thomas Ströhle, and Ruth Breu · 2025
Closest in time.
Self-refinement strategies for llm-based product attribute value extraction
Alexander Brinkmann and Christian Bizer · 2025
Closest in time.
Joyce Cahoon, Prerna Singh, Nick Litombe, Jonathan Larson, Ha Trinh, Yiwen Zhu, Andreas Mueller, Fotis Psallidas, and Carlo Curino · 2025
Closest in time.
Pengfei Cao, Tianyi Men, Wencan Liu, Jingwen Zhang, Xuzhao Li, Xixun Lin, Dianbo Sui, Yanan Cao, Kang Liu, and Jun Zhao · 2025
Closest in time.
Amartya Chakraborty, Paresh Dashore, Nadia Bathaee, Anmol Jain, Anirban Das, Shi-Xiong Zhang, Sambit Sahu, M. Naphade, and Genta Indra Winata · 2025
Closest in time.
Edward Y. Chang and Longling Geng · 2025
Closest in time.
Subhajit Chaudhury, Payel Das, Sarathkrishna Swaminathan, Georgios Kollias, Elliot Nelson, Khushbu Pahwa, Tejaswini Pedapati, Igor Melnyk, and Matthew Riemer · 2025
Closest in time.
Boyu Chen, Zirui Guo, Zidan Yang, Yuluo Chen, Junze Chen, Zhenghao Liu, Chuan Shi, and Cheng Yang · 2025
Closest in time.
A survey on knowledge-oriented retrieval-augmented generation, arXiv preprint arXiv:2503.10677, 2025
Mingyue Cheng, Yucong Luo, Ouyang Jie, Qi Liu, Huijie Liu, Li Li, Shuo Yu, Bohou Zhang, Jiawei Cao, Jie Ma, Daoyu Wang, and Enhong Chen · 2025
Closest in time.
Egor Cherepanov, Nikita Kachaev, A. Kovalev, and Aleksandr I. Panov · 2025
Closest in time.
Prateek Chhikara, Dev Khant, Saket Aryan, Taranjeet Singh, and Deshraj Yadav · 2025
Closest in time.
Llm agents for education: Advances and applications, arXiv preprint arXiv:2503.11733, 2025
Zhendong Chu, Shen Wang, Jian Xie, Tinghui Zhu, Yibo Yan, Jinheng Ye, Aoxiao Zhong, Xuming Hu, Jing Liang, Philip S. Yu, and Qingsong Wen · 2025
Closest in time.
Yung-Sung Chuang, Benjamin Cohen-Wang, Shannon Zejiang Shen, Zhaofeng Wu, Hu Xu, Xi Victoria Lin, James Glass, Shang-Wen Li, and Wen tau Yih · 2025
Closest in time.
Joao Coelho, Jingjie Ning, Jingyuan He, Kangrui Mao, A. Paladugu, Pranav Setlur, Jiahe Jin, James P. Callan, João Magalhães, Bruno Martins, and Chenyan Xiong · 2025
Closest in time.
Injecting knowledge graphs into large language models, arXiv preprint arXiv:2505.07554, 2025
Erica Coppolillo · 2025
Closest in time.
Caia Costello, Simon Guo, Anna Goldie, and Azalia Mirhoseini · 2025
Closest in time.
crewai: Framework for orchestrating role-playing, autonomous ai agents
crewAI Inc · 2025
Closest in time.
Yue Cui, Liuyi Yao, Shuchang Tao, Weijie Shi, Yaliang Li, Bolin Ding, and Xiaofang Zhou · 2025
Closest in time.
Multi-agent collaboration via evolving orchestration, arXiv preprint arXiv:2505.19591, 2025
Yufan Dang, Cheng Qian, Xueheng Luo, Jingru Fan, Zihao Xie, Ruijie Shi, Weize Chen, Cheng Yang, Xiaoyin Che, Ye Tian, Xuantang Xiong, Lei Han, Zhiyuan Liu, and Maosong Sun · 2025
Closest in time.
Knowledge graphs and their reciprocal relationship with large language models
Ramandeep Singh Dehal, Mehak Sharma, and Enayat Rajabi · 2025
Closest in time.
Optimizing human-ai interaction: Innovations in prompt engineering
Rushali Deshmukh, Rutuj Raut, Mayur Bhavsar, Sanika Gurav, and Y. Patil · 2025
Closest in time.
Trail: Trace reasoning and agentic issue localization, arXiv preprint arXiv:2505.08638, 2025
Darshan Deshpande, Varun Gangal, Hersh Mehta, Jitin Krishnan, Anand Kannappan, and Rebecca Qian · 2025
Closest in time.
Frederick Dillon, Gregor Halvorsen, Simon Tattershall, Magnus Rowntree, and Gareth Vanderpool · 2025
Closest in time.
Hanxing Ding, Shuchang Tao, Liang Pang, Zihao Wei, Jinyang Gao, Bolin Ding, Huawei Shen, and Xueqi Chen · 2025
Closest in time.
Reflexive prompt engineering: A framework for responsible prompt engineering and ai interaction design
Christian Djeffal · 2025
Closest in time.
Tool-star: Empowering llm-brained multi-tool reasoner via reinforcement learning
G Dong, Y Chen, X Li, J Jin, H Qian, and Y Zhu… · 2025
Closest in time.
Ehsan Doostmohammadi and Marco Kuhlmann · 2025
Closest in time.
Mom: Linear sequence modeling with mixture-of-memories, arXiv preprint arXiv:2502.13685, 2025
Jusen Du, Weigao Sun, Disen Lan, Jiaxi Hu, and Yu Cheng · 2025
Closest in time.
Eliciting reasoning in language models with cognitive tools, arXiv preprint arXiv:2506.12115, 2025
Brown Ebouky, A. Bartezzaghi, and Mattia Rigotti · 2025
Closest in time.
Abul Ehtesham, Aditi Singh, Gaurav Kumar Gupta, and Saket Kumar · 2025
Closest in time.
Interactive debugging and steering of multi-agent ai systems
Will Epperson, Gagan Bansal, Victor Dibia, Adam Fourney, Jack Gerrits, Erkang Zhu, and Saleema Amershi · 2025
Closest in time.
Gaming tool preferences in agentic llms, arXiv preprint arXiv:2505.18135, 2025
Kazem Faghih, Wenxiao Wang, Yize Cheng, Siddhant Bharti, Gaurang Sriramanan, S. Balasubramanian, Parsa Hosseini, and S. Feizi · 2025
Closest in time.
Siqi Fan, Xiusheng Huang, Yiqun Yao, Xuezhi Fang, Kang Liu, Peng Han, Shuo Shang, Aixin Sun, and Yequan Wang · 2025
Closest in time.
Junfeng Fang, Zijun Yao, Ruipeng Wang, Haokai Ma, Xiang Wang, and Tat-Seng Chua · 2025
Closest in time.
George Fatouros, Georgios Makridis, George Kousiouris, John Soldatos, A. Tsadimas, and D. Kyriazis · 2025
Closest in time.
Mcp-zero: Proactive toolchain construction for llm agents from scratch
Xiang Fei, Xiawu Zheng, and Hao Feng · 2025
Closest in time.
Get experience from practice: Llm agents with record & replay, arXiv preprint arXiv:2505.17716, 2025
Erhu Feng, Wenbo Zhou, Zibin Liu, Le Chen, Yunpeng Dong, Cheng Zhang, Yisheng Zhao, Dong Du, Zhi-Hua Zhou, Yubin Xia, and Haibo Chen · 2025
Closest in time.
M. Ferrag, Norbert Tihanyi, and M. Debbah · 2025
Closest in time.
Brainvis: Exploring the bridge between brain and visual signals via image reconstruction
Honghao Fu, Hao Wang, Jing Jih Chin, and Zhiqi Shen · 2025
Closest in time.
Rag-mcp: Mitigating prompt bloat in llm tool selection via retrieval-augmented generation
Tiantian Gan and Qiyao Sun · 2025
Closest in time.
Mark: Memory augmented refinement of knowledge, arXiv preprint arXiv:2505.05177, 2025
Anish Ganguli, Prabal Deb, and Debleena Banerjee · 2025
Closest in time.
Chen Gao, Xiaochong Lan, Zhihong Lu, Jinzhu Mao, Jinghua Piao, Huandong Wang, Depeng Jin, and Yong Li · 2025
Closest in time.
Innate reasoning is not enough: In-context learning enhances reasoning large language models with less overthinking
Yuyao Ge, Shenghua Liu, Yiwei Wang, Lingrui Mei, Lizhe Chen, Baolong Bi, and Xueqi Cheng · 2025
Closest in time.
Innovators and transformers: enhancing supply chain employee training with an innovative application of a large language model
Arda Gezdur and J. Bhattacharjya · 2025
Closest in time.
Hierarchical lexical graph for enhanced multi-hop retrieval, arXiv preprint arXiv:2506.08074, 2025
Abdellah Ghassel, Ian Robinson, Gabriel Tanase, Hal Cooper, Bryan Thompson, Zhen Han, V. Ioannidis, Soji Adeshina, and H. Rangwala · 2025
Closest in time.
Ekaterina Grishina, Mikhail Gorbunov, and Maxim Rakhuba · 2025
Closest in time.
Hdtcnet: A hybrid-dimensional convolutional network for multivariate time series classification
Yongli Gu, Xiang Yan, Hanlin Qin, Naveed Akhtar, Shuai Yuan, Honghao Fu, Shuowen Yang, and Ajmal Mian · 2025
Closest in time.
Shengyue Guan, Haoyi Xiong, Jindong Wang, Jiang Bian, Bin Zhu, and Jian guang Lou · 2025
Closest in time.
G1: Teaching llms to reason on graphs with reinforcement learning
Xiaojun Guo, Ang Li, Yifei Wang, Stefanie Jegelka, and Yisen Wang · 2025
Closest in time.
Idan Habler, Ken Huang, Vineeth Sai Narajala, and Prashant Kulkarni · 2025
Closest in time.
John Halloran · 2025
Closest in time.
Feijiang Han, Licheng Guo, Hengtao Cui, and Zhiyuan Lyu · 2025
Closest in time.
Evaluating the sensitivity of llms to prior context, arXiv preprint arXiv:2506.00069, 2025
R. Hankache, Kingsley Nketia Acheampong, Liang Song, Marek Brynda, Raad Khraishi, and Greig A. Cowan · 2025
Closest in time.
Mohanakrishnan Hariharan · 2025
Closest in time.
Kostas Hatalis, Despina Christou, and Vyshnavi Kondapalli · 2025
Closest in time.
Jacky He, Guiran Liu, Binrong Zhu, Hanlu Zhang, Hongye Zheng, and Xiaokai Wang · 2025
Closest in time.
Tooraj Helmi · 2025
Closest in time.
Enhancing memory retrieval in generative agents through llm-trained cross attention networks
Chuanyang Hong and Qingyun He · 2025
Closest in time.
Fg-rag: Enhancing query-focused summarization with context-aware fine-grained graph rag
Yubin Hong, Chaofan Li, Jingyi Zhang, and Yingxia Shao · 2025
Closest in time.
Xinyi Hou, Yanjie Zhao, Shenao Wang, and Haoyu Wang · 2025
Closest in time.
Junhao Hu, Wenrui Huang, Weidong Wang, Zhenwen Li, Tiancheng Hu, Zhixia Liu, XuSheng Chen, Tao Xie, and Yizhou Shan · 2025
Closest in time.
Chengkai Huang, Hongtao Huang, Tong Yu, Kaige Xie, Junda Wu, Shuai Zhang, Julian J. McAuley, Dietmar Jannach, and Lina Yao · 2025
Closest in time.
Episodic memories generation and evaluation benchmark for large language models
Alexis Huet, Zied Ben-Houidi, and Dario Rossi · 2025
Closest in time.
What is agent communication protocol (acp)?
IBM · 2025
Closest in time.
Jace.ai web agent, 2024
Jace.AI · 2025
Closest in time.
Tejas Jade and Alex Yartsev · 2025
Closest in time.
Cheonsu Jeong · 2025
Closest in time.
Evaluating large language model with knowledge oriented language specific simple question answering
Bowen Jiang, Runchuan Zhu, Jiang Wu, Zinco Jiang, Yifan He, Junyuan Gao, Jia Yu, Rui Min, Yinfan Wang, Haote Yang, et al · 2025
Closest in time.
A comprehensive survey on multi-agent cooperative decision-making: Scenarios, approaches, challenges and perspectives
Weiqiang Jin, Hongyang Du, Biao Zhao, Xingwu Tian, Bohang Shi, and Guang Yang · 2025
Closest in time.
Kurmanbek Kaiyrbekov, Nic Dobbins, and Sean D. Mooney · 2025
Closest in time.
Eser Kandogan, Nikita Bhutani, Dan Zhang, Rafael Li Chen, Sairam Gurajada, and Estevam R. Hruschka · 2025
Closest in time.
Memory os of ai agent, arXiv preprint arXiv:2506.06326, 2025
Jiazheng Kang, Mingming Ji, Zhe Zhao, and Ting Bai · 2025
Closest in time.
Kiran Kate, Tejaswini Pedapati, Kinjal Basu, Yara Rizk, Vijil Chenthamarakshan, Subhajit Chaudhury, Mayank Agarwal, and Ibrahim Abdelaziz · 2025
Closest in time.
Richard Katrix, Quentin Carroway, Rowan Hawkesbury, and Matthias Heathfield · 2025
Closest in time.
Cdf-rag: Causal dynamic feedback for adaptive retrieval-augmented generation
Elahe Khatibi, Ziyu Wang, and Amir M. Rahmani · 2025
Closest in time.
Jiin Kim, Byeongjun Shin, Jinha Chung, and Minsoo Rhu · 2025
Closest in time.
Discovering multi-agent systems for resource-centric business process simulation
Lukas Kirchdorfer, Robert Blümel, T. Kampik, Han van der Aa, and Heiner Stuckenschmidt · 2025
Closest in time.
Andrew Kiruluta, Preethi Raju, and Priscilla Burity · 2025
Closest in time.
Vincent Koc, Jacques Verre, Douglas Blank, and Abigail Morgan · 2025
Closest in time.
Cognitive prompts using guilford’s structure of intellect model
Oliver Kramer · 2025
Closest in time.
L. Krupp, Daniel Geissler, P. Lukowicz, and Jakob Karolus · 2025
Closest in time.
Detecting and mitigating bias in llms through knowledge graph-augmented training
Rajeev Kumar, Harishankar Kumar, and Kumari Shalini · 2025
Closest in time.
Taeyoon Kwon, Dongwook Choi, Sunghwan Kim, Hyojun Kim, Seungjun Moon, Beong woo Kwak, Kuan-Hao Huang, and Jinyoung Yeo · 2025
Closest in time.
Xiaochong Lan, Jie Feng, Jia Lei, Xinlei Shi, and Yong Li · 2025
Closest in time.
Memory in langgraph
LangChain Team · 2025
Closest in time.
Dohyun Lee, Seungil Chad Lee, Chanwoo Yang, Yujin Baek, and Jaegul Choo · 2025
Closest in time.
Yiming Lei, Zhizheng Yang, Zeming Liu, Haitao Leng, Shaoguo Liu, Tingting Gao, Qingjie Liu, and Yunhong Wang · 2025
Closest in time.
Cort: Code-integrated reasoning within thinking, arXiv preprint arXiv:2506.09820, 2025
Chengpeng Li, Zhengyang Tang, Ziniu Li, Mingfeng Xue, Keqin Bao, Tian Ding, Ruoyu Sun, Benyou Wang, Xiang Wang, Junyang Lin, and Dayiheng Liu · 2025
Closest in time.
Qiaomu Li and Ying Xie · 2025
Closest in time.
Drs: Deep question reformulation with structured output
Zhecheng Li, Yiwei Wang, Bryan Hooi, Yujun Cai, Nanyun Peng, and Kai-Wei Chang · 2025
Closest in time.
Vulnerability of llms to vertically aligned text manipulations
Zhecheng Li, Yiwei Wang, Bryan Hooi, Yujun Cai, Zhen Xiong, Nanyun Peng, and Kai-Wei Chang · 2025
Closest in time.
Texture or semantics? vision-language models get lost in font recognition
Zhecheng Li, Guoxian Song, Yujun Cai, Zhen Xiong, Junsong Yuan, and Yiwei Wang · 2025
Closest in time.
Guannan Liang and Qianqian Tong · 2025
Closest in time.
Jintao Liang, Gang Su, Huifeng Lin, You Wu, Rui Zhao, and Ziyue Li · 2025
Closest in time.
Xiaoxuan Liao, Binrong Zhu, Jacky He, Guiran Liu, Hongye Zheng, and Jia Gao · 2025
Closest in time.
Simulating macroeconomic expectations using llm agents, arXiv preprint arXiv:2505.17648, 2025
Jianhao Lin, Lexuan Sun, and Yixin Yan · 2025
Closest in time.
Guangyi Liu, Pengxiang Zhao, Liang Liu, Yaxuan Guo, Han Xiao, Weifeng Lin, Yuxiang Chai, Yue Han, Shuai Ren, Hao Wang, Xiaoyu Liang, Wenhao Wang, Tianze Wu, Linghao Li, Guanjing Xiong, Yong Liu, and Hongsheng Li · 2025
Closest in time.
Joseph R. Loffredo and Suyeol Yun · 2025
Closest in time.
Junting Lu, Zhiyang Zhang, Fangkai Yang, Jue Zhang, Lu Wang, Chao Du, Qingwei Lin, Saravan Rajmohan, Dongmei Zhang, and Qi Zhang · 2025
Closest in time.
Elias Lumer, Anmol Gulati, V. K. Subbiah, Pradeep Honaganahalli Basavaraju, and James A. Burke · 2025
Closest in time.
Feng Luo, Yu-Neng Chuang, Guanchu Wang, Hoang Anh Duy Le, Shaochen Zhong, Hongyi Liu, Jiayi Yuan, Yang Sui, Vladimir Braverman, Vipin Chaudhary, and Xia Hu · 2025
Closest in time.
Panagiotis Lymperopoulos and Vasanth Sarathy · 2025
Closest in time.
Deepshop: A benchmark for deep research shopping agents, arXiv preprint arXiv:2506.02839, 2025
Yougang Lyu, Xiaoyu Zhang, Lingyong Yan, M. D. Rijke, Zhaochun Ren, and Xiuying Chen · 2025
Closest in time.
Automated creation of reusable and diverse toolsets for enhancing llm reasoning
Zhiyuan Ma, Zhenya Huang, Jiayu Liu, Minmao Wang, Hongke Zhao, and Xin Li · 2025
Closest in time.
Xinji Mai, Haotian Xu, W. Xing, Weinong Wang, Yingying Zhang, and Wenqiang Zhang · 2025
Closest in time.
Compass-v2 technical report, arXiv preprint arXiv:2504.15527, 2025
Sophia Maria · 2025
Closest in time.
Vasilije Markovic, Lazar Obradović, L’aszl’o Hajdu, and Jovan Pavlović · 2025
Closest in time.
Towards enterprise-ready computer using generalist agent
Sami Marreed, Alon Oved, Avi Yaeli, Segev Shlomov, Ido Levy, Offer Akrabi, Aviad Sela, Asaf Adi, and Nir Mashkif · 2025
Closest in time.
Latent multi-head attention for small language models, arXiv preprint arXiv:2506.09342, 2025
Sushant Mehta, R. Dandekar, R. Dandekar, and S. Panat · 2025
Closest in time.
Lingrui Mei, Shenghua Liu, Yiwei Wang, Baolong Bi, Yuyao Ge, Jun Wan, Yurong Wu, and Xueqi Cheng · 2025
Closest in time.
Yapeng Mi, Zhi Gao, Xiaojian Ma, and Qing Li · 2025
Closest in time.
Pel, a programming language for orchestrating ai agents, arXiv preprint arXiv:2505.13453, 2025
Behnam Mohammadi · 2025
Closest in time.
Manisha Mukherjee, Sungchul Kim, Xiang Chen, Dan Luo, Tong Yu, and Tung Mai · 2025
Closest in time.
Tergel Munkhbat, Namgyu Ho, Seohyun Kim, Yongjin Yang, Yujin Kim, and Se young Yun · 2025
Closest in time.
Modp: Multi objective directional prompting, arXiv preprint arXiv:2504.18722, 2025
Aashutosh Nema, Samaksh Gulati, Evangelos Giakoumakis, and Bipana Thapaliya · 2025
Closest in time.
A grounded memory system for smart personal assistants, arXiv preprint arXiv:2505.06328, 2025
Felix Ocker, J. Deigmöller, Pavel Smirnov, and Julian Eggert · 2025
Closest in time.
Computer-using agent, 2025
OpenAI · 2025
Closest in time.
Swarm: Educational framework exploring ergonomic, lightweight multi-agent orchestration
OpenAI · 2025
Closest in time.
Jonas Oppenlaender · 2025
Closest in time.
Qianjun Pan, Wenkai Ji, Yuyang Ding, Junsong Li, Shilian Chen, Junyi Wang, Jie Zhou, Qin Chen, Min Zhang, Yulan Wu, and Liang He · 2025
Closest in time.
Bo Pang, Hanze Dong, Jiacheng Xu, Silvio Savarese, Yingbo Zhou, and Caiming Xiong · 2025
Closest in time.
Soya Park, J. Zamfirescu-Pereira, and Chinmay Kulkarni · 2025
Closest in time.
The berkeley function calling leaderboard (bfcl): From tool use to agentic evaluation of large language models
Shishir G. Patil, Huanzhi Mao, Charlie Cheng-Jie Ji, Fanjia Yan, Vishnu Suresh, Ion Stoica, and Joseph E. Gonzalez · 2025
Closest in time.
Shuva Paul, Farhad Alemi, and Richard Macwan · 2025
Closest in time.
Advancing feature extraction in healthcare through the integration of knowledge graphs and large language models
Fahmida Liza Piya and Rahmatollah Beheshti · 2025
Closest in time.
Agentic large language models, a survey, arXiv preprint arXiv:2503.23037, 2025
A. Plaat, M. V. Duijn, N. V. Stein, Mike Preuss, P. V. D. Putten, and K. Batenburg · 2025
Closest in time.
Teaching llms music theory with in-context learning and chain-of-thought prompting: Pedagogical strategies for machines
Liam Pond and Ichiro Fujinaga · 2025
Closest in time.
Toolrl: Reward is all tool learning needs, arXiv preprint arXiv:2504.13958, 2025
Cheng Qian, Emre Can Acikgoz, Qi He, Hongru Wang, Xiusi Chen, Dilek Hakkani-Tur, Gokhan Tur, and Heng Ji · 2025
Closest in time.
Changze Qiao and Mingming Lu · 2025
Closest in time.
Jiahao Qiu, Xinzhe Juan, Yiming Wang, Ling Yang, Xuan Qi, Tongcheng Zhang, Jiacheng Guo, Yifu Lu, Zixin Yao, Hongru Wang, Shilong Liu, Xun Jiang, Liu Leqi, and Mengdi Wang · 2025
Closest in time.
Xiaoye Qu, Yafu Li, Zhao yu Su, Weigao Sun, Jianhao Yan, Dongrui Liu, Ganqu Cui, Daizong Liu, Shuxian Liang, Junxian He, Peng Li, Wei Wei, Jing Shao, Chaochao Lu, Yue Zhang, Xian-Sheng Hua, Bowen Zhou, and Yu Cheng · 2025
Closest in time.
On the robustness of agentic function calling
Ella Rabinovich and Ateret Anaby-Tavor · 2025
Closest in time.
Shaina Raza, Ranjan Sapkota, Manoj Karkee, and Christos Emmanouilidis · 2025
Closest in time.
Shuo Ren, Pu Jian, Zhenjiang Ren, Chunlin Leng, Can Xie, and Jiajun Zhang · 2025
Closest in time.
Juan David Salazar Rodriguez, Sam Conrad Joyce, and Julfendi Julfendi · 2025
Closest in time.
When2call: When (not) to call tools
Hayley Ross, A. Mahabaleshwarkar, and Yoshi Suhara · 2025
Closest in time.
J. Rosser and Jakob N. Foerster · 2025
Closest in time.
Meminsight: Autonomous memory augmentation for llm agents, arXiv preprint arXiv:2503.21760, 2025
Rana Salama, Jason Cai, Michelle Yuan, Anna Currey, Monica Sunkara, Yi Zhang, and Yassine Benajiba · 2025
Closest in time.
Alaa Saleh, Sasu Tarkoma, Praveen Kumar Donta, Naser Hossein Motlagh, S. Dustdar, Susanna Pirttikangas, and Lauri Lov’en · 2025
Closest in time.
Model context protocol: Enhancing llm performance for observability and analytics
Narendra Reddy Sanikommu · 2025
Closest in time.
Diverse prompts: Illuminating the prompt space of large language models with map-elites
G. Santos, Rita Maria Silva Julia, and Marcelo Zanchetta do Nascimento · 2025
Closest in time.
Ranjan Sapkota, Konstantinos I. Roumeliotis, and Manoj Karkee · 2025
Closest in time.
Anjana Sarkar and Soumyendu Sarkar · 2025
Closest in time.
Collex - a multimodal agentic rag system enabling interactive exploration of scientific collections
Florian Schneider, Narges Baba Ahmadi, Niloufar Baba Ahmadi, Iris Vogel, Martin Semmann, and Christian Biemann · 2025
Closest in time.
The evolving landscape of llm- and vlm-integrated reinforcement learning
Sheila Schoepp, Masoud Jafaripour, Yingyue Cao, Tianpei Yang, Fatemeh Abdollahi, Shadan Golestan, Zahin Sufiyan, Osmar R. Zaiane, and Matthew E. Taylor · 2025
Closest in time.
Cognitive memory in large language models, arXiv preprint arXiv:2504.02441, 2025
Lianlei Shan, Shixian Luo, Zezhou Zhu, Yu Yuan, and Yong Wu · 2025
Closest in time.
Junhong Shen, Hao Bai, Lunjun Zhang, Yifei Zhou, Amrith Setlur, Shengbang Tong, Diego Caples, Nan Jiang, Tong Zhang, Ameet Talwalkar, and Aviral Kumar · 2025
Closest in time.
Taskcraft: Automated generation of agentic tasks, arXiv preprint arXiv:2506.10055, 2025
Dingfeng Shi, Jingyi Cao, Qianben Chen, Weichen Sun, Weizhen Li, Hongxuan Lu, Fangchen Dong, Tianrui Qin, King Zhu, Minghao Liu, Jian Yang, Ge Zhang, Jiaheng Liu, Changwang Zhang, Jun Wang, Y. Jiang, and Wangchunshu Zhou · 2025
Closest in time.
Altun Shukurlu · 2025
Closest in time.
Joykirat Singh, Raghav Magazine, Yash Pandya, and A. Nambi · 2025
Closest in time.
Aarush Sinha and CU Omkumar · 2025
Closest in time.
Colin Sisate, Alistair Goldfinch, Vincent Waterstone, Sebastian Kingsley, and Mariana Blackthorn · 2025
Closest in time.
Efficient document retrieval with g-retriever
Manthankumar Solanki · 2025
Closest in time.
A comprehensive survey on integrating large language models with knowledge-based methods
Lilian Some, Wenli Yang, Michael Bain, and Byeong Kang · 2025
Closest in time.
R1-searcher: Incentivizing the search capability in llms via reinforcement learning
Huatong Song, Jinhao Jiang, Yingqian Min, Jie Chen, Zhipeng Chen, Wayne Xin Zhao, Lei Fang, and Ji-Rong Wen · 2025
Closest in time.
Conversational alignment with artificial intelligence in context
R. Sterken and James Ravi Kirkpatrick · 2025
Closest in time.
Learn-by-interact: A data-centric framework for self-adaptive agents in realistic environments
Hongjin Su, Ruoxi Sun, Jinsung Yoon, Pengcheng Yin, Tao Yu, and Sercan Ö. Arık · 2025
Closest in time.
Yang Sui, Yu-Neng Chuang, Guanchu Wang, Jiamu Zhang, Tianyi Zhang, Jiayi Yuan, Hongyi Liu, Andrew Wen, Shaochen Zhong, Hanjie Chen, and Xia Hu · 2025
Closest in time.
Lijun Sun, Yijun Yang, Qiqi Duan, Yuhui Shi, Chao Lyu, Yu-Cheng Chang, Chin-Teng Lin, and Yang Shen · 2025
Closest in time.
Announcing the agent2agent protocol (a2a)
Rao Surapaneni, Miku Jha, Michael Vakoc, and Todd Segal · 2025
Closest in time.
Daniel Szelogowski · 2025
Closest in time.
Training a generally curious agent, arXiv preprint arXiv:2502.17543, 2025
Fahim Tajwar, Yiding Jiang, Abitha Thankaraj, Sumaita Sadia Rahman, J. Z. Kolter, Jeff Schneider, and Ruslan Salakhutdinov · 2025
Closest in time.
K. Tallam · 2025
Closest in time.
A survey on (m)llm-based gui agents, arXiv preprint arXiv:2504.13865, 2025
Fei Tang, Haolei Xu, Hang Zhang, Siqi Chen, Xingyu Wu, Yongliang Shen, Wenqi Zhang, Guiyang Hou, Zeqi Tan, Yuchen Yan, Kaitao Song, Jian Shao, Weiming Lu, Jun Xiao, and Yueting Zhuang · 2025
Closest in time.
Yao Tao, Yehui Tang, Yun Wang, Mingjian Zhu, Hailin Hu, and Yunhe Wang · 2025
Closest in time.
Typhoon t1: An open thai reasoning model, arXiv preprint arXiv:2502.09042, 2025
Pittawat Taveekitworachai, Potsawee Manakul, Kasima Tharnpipitchai, and Kunat Pipatanakul · 2025
Closest in time.
The future of ai: From parameter scaling to context scaling
36Kr Editorial Team · 2025
Closest in time.
Ego-r1: Chain-of-tool-thought for ultra-long egocentric video reasoning
S Tian, R Wang, H Guo, P Wu, Y Dong, and X Wang… · 2025
Closest in time.
Despina Tomkou, George Fatouros, Andreas Andreou, Georgios Makridis, F. Liarokapis, Dimitrios Dardanis, Athanasios Kiourtis, John Soldatos, and D. Kyriazis · 2025
Closest in time.
Llm-based text style transfer: Have we taken a step forward?
Martina Toshevska and Sonja Gievska · 2025
Closest in time.
Multi-agent collaboration mechanisms: A survey of llms, arXiv preprint arXiv:2501.06322, 2025
Khanh-Tung Tran, Dung Dao, Minh-Duong Nguyen, Quoc-Viet Pham, Barry O’Sullivan, and Hoang D. Nguyen · 2025
Closest in time.
Multi-agent systems execute arbitrary malicious code, arXiv preprint arXiv:2503.12188, 2025
Harold Triedman, Rishi Jha, and Vitaly Shmatikov · 2025
Closest in time.
Towards conversational diagnostic artificial intelligence
Tao Tu, M. Schaekermann, Anil Palepu, Khaled Saab, Jan Freyberg, Ryutaro Tanno, Amy Wang, Brenna Li, Mohamed Amin, Yong Cheng, Elahe Vedadi, Nenad Tomašev, Shekoofeh Azizi, Karan Singhal, Le Hou, Albert Webson, Kavita Kulkarni, S. Mahdavi, Christopher Semturs, Juraj Gottweis, Joelle Barral, Katherine Chou, Greg S. Corrado, Yossi Matias, A. Karthikesalingam, and Vivek Natarajan · 2025
Closest in time.
D-cipher: Dynamic collaborative intelligent multi-agent system with planner and heterogeneous executors for offensive security
Meet Udeshi, Minghao Shao, Haoran Xi, Nanda Rani, Kimberly Milner, Venkata Sai Charan Putrevu, Brendan Dolan-Gavitt, S. K. Shukla, P. Krishnamurthy, F. Khorrami, Ramesh Karri, and Muhammad Shafique · 2025
Closest in time.
Saeid Ario Vaghefi, Aymane Hachcham, Veronica Grasso, Jiska Manicus, Nakiete Msemo, C. Senni, and Markus Leippold · 2025
Closest in time.
From symbolic to neural and back: Exploring knowledge graph-large language model synergies
Blavz vSkrlj, Boshko Koloski, S. Pollak, and Nada Lavravc · 2025
Closest in time.
Jun Wan and Lingrui Mei · 2025
Closest in time.
Luanbo Wan and Weizhi Ma · 2025
Closest in time.
Bing Wang, Xinnian Liang, Jian Yang, Hui Huang, Shuangzhi Wu, Peihao Wu, Lu Lu, Zejun Ma, and Zhoujun Li · 2025
Closest in time.
Jingjin Wang · 2025
Closest in time.
Yingming Wang and Pepa Atanasova · 2025
Closest in time.
Webagent-r1: Training web agents via end-to-end multi-turn reinforcement learning
Zhepei Wei, Wenlin Yao, Yao Liu, Weizhi Zhang, Qin Lu, Liang Qiu, Changlong Yu, Puyang Xu, Chao Zhang, Bing Yin, Hyokun Yun, and Lihong Li · 2025
Closest in time.
Caim: Development and evaluation of a cognitive ai memory framework for long-term interaction with intelligent agents
Rebecca Westhäußer, Frederik Berenz, Wolfgang Minker, and Sebastian Zepf · 2025
Closest in time.
Agent communications language — Wikipedia, the free encyclopedia, 2025
Wikipedia contributors · 2025
Closest in time.
Beong woo Kwak, Minju Kim, Dongha Lim, Hyungjoo Chae, Dongjin Kang, Sunghwan Kim, Dongil Yang, and Jinyoung Yeo · 2025
Closest in time.
Webdancer: Towards autonomous information seeking agency, arXiv preprint arXiv:2505.22648, 2025
Jialong Wu, Baixuan Li, Runnan Fang, Wenbiao Yin, Liwen Zhang, Zhengwei Tao, Dingchu Zhang, Zekun Xi, Yong Jiang, Pengjun Xie, Fei Huang, and Jingren Zhou · 2025
Closest in time.
Zengqing Wu and Takayuki Ito · 2025
Closest in time.
Menglin Xia, Victor Ruehle, Saravan Rajmohan, and Reza Shokri · 2025
Closest in time.
Zhishang Xiang, Chuanjie Wu, Qinggang Zhang, Shengyuan Chen, Zijin Hong, Xiao Huang, and Jinsong Su · 2025
Closest in time.
Minicpm4: Ultra-efficient llms on end devices, arXiv preprint arXiv:2506.07900, 2025
MiniCPM Team Chaojun Xiao, Yuxuan Li, Xu Han, Yuzhuo Bai, Jie Cai, Haotian Chen, Wentong Chen, Xin Cong, Ganqu Cui, Ning Ding, Shengda Fan, Yewei Fang, Zixuan Fu, Wenyu Guan, Yitong Guan, Junshao Guo, Yu-Xuan Han, Bingxiang He, Yuxian Huang, Cunliang Kong, Qiu-Tong Li, Siyuan Li, Wenhao Li, Yanghao Li, Yishan Li, Zhen Li, Dan Liu, Biyuan Lin, Yankai Lin, Xiang Long, Quanyu Lu, Ya-Ting Lu, Pei Luo, Hongya Lyu, Litu Ou, Yinxu Pan, Zekai Qu, Qundong Shi, Zijun Song, Jiayu Su, Zhou Su, Ao Sun, Xiang ping Sun, Peijun Tang, Fang-Ming Wang, Feng Wang, Shuo Wang, Yudong Wang, Yesai Wu, Zhenyu Xiao, Jie Xie, Zi-Kang Xie, Yukun Yan, Jia-Li Yuan, Kai Zhang, Lei Zhang, Linyu Zhang, Xueren Zhang, Yudi Zhang, Hengyu Zhao, Weilin Zhao, Weilun Zhao, Yuanqian Zhao, Zhijun Zheng, Ge Zhou, Jie Zhou, Wei Zhou, Zihan Zhou, Zi-An Zhou, Zhiyuan Liu, Guoyang Zeng, Chaochao Jia, Dahai Li, and Maosong Sun · 2025
Closest in time.
Yue Xing, Tao Yang, Yijiashun Qi, Minggu Wei, Yu Cheng, and Honghui Xin · 2025
Closest in time.
Rag-gym: Systematic optimization of language agents for retrieval-augmented generation
Guangzhi Xiong, Qiao Jin, Xiao Wang, Yin Fang, Haolin Liu, Yifan Yang, Fangyuan Chen, Zhixing Song, Dengyu Wang, Minjia Zhang, Zhiyong Lu, and Aidong Zhang · 2025
Closest in time.
Redstar: Does scaling long-cot data unlock better slow-reasoning systems?
Haotian Xu, Xing Wu, Weinong Wang, Zhongzhi Li, Da Zheng, Boyuan Chen, Yi Hu, Shijia Kang, Jiaming Ji, Yingying Zhang, et al · 2025
Closest in time.
Shuhang Xu and Fangwei Zhong · 2025
Closest in time.
A survey of attacks on large language models, arXiv preprint arXiv:2505.12567, 2025
Wenrui Xu and Keshab K. Parhi · 2025
Closest in time.
Eric Xue, Ke Chen, Zeyi Huang, Yuyang Ji, Yong Jae Lee, and Haohan Wang · 2025
Closest in time.
Beyond self-talk: A communication-centric survey of llm-based multi-agent systems
Bingyu Yan, Xiaoming Zhang, Litian Zhang, Lian Zhang, Ziyi Zhou, Dezhuang Miao, and Chaozhuo Li · 2025
Closest in time.
Mmada: Multimodal large diffusion language models, arXiv preprint arXiv:2505.15809v1, 2025
Ling Yang, Ye Tian, Bowen Li, Xinchen Zhang, Ke Shen, Yunhai Tong, and Mengdi Wang · 2025
Closest in time.
Who is in the spotlight: The hidden bias undermining multimodal retrieval-augmented generation
Jiayu Yao, Shenghua Liu, Yiwei Wang, Lingrui Mei, Baolong Bi, Yuyao Ge, Zhecheng Li, and Xueqi Cheng · 2025
Closest in time.
Junjie Ye, Zhengyin Du, Xuesong Yao, Weijian Lin, Yufei Xu, Zehui Chen, Zaiyuan Wang, Sining Zhu, Zhiheng Xi, Siyu Yuan, Tao Gui, Qi Zhang, Xuanjing Huang, and Jiechao Chen · 2025
Closest in time.
Survey on evaluation of llm-based agents, arXiv preprint arXiv:2503.16416, 2025
Asaf Yehudai, Lilach Eden, Alan Li, Guy Uziel, Yilun Zhao, Roy Bar-Haim, Arman Cohan, and Michal Shmueli-Scheuer · 2025
Closest in time.
Irony detection, reasoning and understanding in zero-shot learning
Peiling Yi and Yuhan Xia · 2025
Closest in time.
Fan Yin, Zifeng Wang, I-Hung Hsu, Jun Yan, Ke Jiang, Yanfei Chen, Jindong Gu, Long T. Le, Kai-Wei Chang, Chen-Yu Lee, Hamid Palangi, and Tomas Pfister · 2025
Closest in time.
Miao Yu, Fanci Meng, Xinyun Zhou, Shilong Wang, Junyuan Mao, Linsey Pang, Tianlong Chen, Kun Wang, Xinfeng Li, Yongfeng Zhang, Bo An, and Qingsong Wen · 2025
Closest in time.
Zhao yu Su, Linjie Li, Mingyang Song, Yunzhuo Hao, Zhengyuan Yang, Jun Zhang, Guanjie Chen, Jiawei Gu, Juntao Li, Xiaoye Qu, and Yu Cheng · 2025
Closest in time.
Agent-r: Training language model agents to reflect via iterative self-training
Siyu Yuan, Zehui Chen, Zhiheng Xi, Junjie Ye, Zhengyin Du, and Jiecao Chen · 2025
Closest in time.
Murong Yue · 2025
Closest in time.
Yirong Zeng, Xiao Ding, Yuxian Wang, Weiwen Liu, Wu Ning, Yutai Hou, Xu Huang, Bing Qin, and Ting Liu · 2025
Closest in time.
Chaoyun Zhang, He Huang, Chiming Ni, Jian Mu, Si Qin, Shilin He, Lu Wang, Fangkai Yang, Pu Zhao, Chao Du, et al · 2025
Closest in time.
Qi Zhao, Hongyu Yang, Qi Song, Xinwei Yao, and Xiangyang Li · 2025
Closest in time.
Junhao Zheng, Xidi Cai, Qiuke Li, Duzhen Zhang, Zhongzhi Li, Yingying Zhang, Le Song, and Qianli Ma · 2025
Closest in time.
Trustrag: Enhancing robustness and trustworthiness in rag, arXiv preprint arXiv:2501.00879, 2025
Huichi Zhou, Kin-Hei Lee, Zhonghao Zhan, Yue Chen, Zhenhao Li, Zhaoyang Wang, Hamed Haddadi, and Emine Yilmaz · 2025
Closest in time.
Jiachen Zhu, Menghui Zhu, Renting Rui, Rong Shan, Congmin Zheng, Bo Chen, Yunjia Xi, Jianghao Lin, Weiwen Liu, Ruiming Tang, Yong Yu, and Weinan Zhang · 2025
Closest in time.
Zhixiong Zhuang, Maria-Irina Nicolae, Hui-Po Wang, and Mario Fritz · 2025
Closest in time.
A survey on large language model based human-agent systems
Henry Peng Zou, Wei-Chieh Huang, Yaozu Wu, Yankai Chen, Chunyu Miao, Hoang Nguyen, Yue Zhou, Weizhi Zhang, Liancheng Fang, Langzhou He, Yangning Li, Yuwei Cao, Dongyuan Li, Renhe Jiang, and Philip S. Yu · 2025
Closest in time.
Self-adapting language models, arXiv preprint arXiv:2506.10943, 2025
Adam Zweiger, Jyothish Pari, Han Guo, Ekin Akyürek, Yoon Kim, and Pulkit Agrawal · 2025
Closest in time.