Evolving through the looking glass: Learning improved search spaces with variational autoencoders. In International Conference on Parallel Problem Solving from Nature . Springer, 371–384
Peter J Bentley, Soo Ling Lim, Adam Gaier, and Linh Tran. 2022 · 2022
Later among the works it cites.
Tweetnlp: Cutting-edge natural language processing for social media
Original
Jose Camacho-Collados, Kiamehr Rezaee, Talayeh Riahi, Asahi Ushio, Daniel Loureiro, Dimosthenis Antypas, Joanne Boisson, Luis Espinosa-Anke, Fangyu Liu, Eugenio Martínez-Cámara, et al · 2022
Later among the works it cites.
Data distributional properties drive emergent few-shot learning in transformers. In Proceedings of the Conference on Neural Information Processing Systems (NeurIPS)
Stephanie C. Y. Chan, Adam Santoro, Andrew K Lampinen, Jane X Wang, Aaditya Singh, Pierre H Richemond, Jay McClelland, and Felix Hill. 2022 · 2022
Later among the works it cites.
Scaling instruction-finetuned language models
Original
Hyung Won Chung, Le Hou, Shayne Longpre, Barret Zoph, Yi Tay, William Fedus, Yunxuan Li, Xuezhi Wang, Mostafa Dehghani, Siddhartha Brahma, et al · 2022
Later among the works it cites.
Deep learning for text style transfer: A survey
Di Jin, Zhijing Jin, Zhiting Hu, Olga Vechtomova, and Rada Mihalcea. 2022 · 2022
Later among the works it cites.
End-to-end Symbolic Regression with Transformers. In Advances in Neural Information Processing Systems
Pierre-Alexandre Kamienny, Stéphane d’Ascoli, Guillaume Lample, and Francois Charton. 2022 · 2022
Later among the works it cites.
Mutation Models: Learning to Generate Levels by Imitating Evolution. In Proceedings of the 17th International Conference on the Foundations of Digital Games . 1–9
Ahmed Khalifa, Julian Togelius, and Michael Cerny Green. 2022 · 2022
Later among the works it cites.
Competition-level Code Generation with Alphacode
Yujia Li, David Choi, Junyoung Chung, Nate Kushman, Julian Schrittwieser, Rémi Leblond, Tom Eccles, James Keeling, Felix Gimeno, Agustin Dal Lago, et al · 2022
Later among the works it cites.
Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering. In NeurIPS
Pan Lu, Swaroop Mishra, Tony Xia, Liang Qiu, Kai-Wei Chang, Song-Chun Zhu, Oyvind Tafjord, Peter Clark, and Ashwin Kalyan. 2022 · 2022
Later among the works it cites.
Simple genetic operators are universal approximators of probability distributions (and other advantages of expressive encodings). In Proceedings of the Genetic and Evolutionary Computation Conference (GECCO) . 739–748
Elliot Meyerson, Xin Qiu, and Risto Miikkulainen. 2022 · 2022
Later among the works it cites.
Codegen: An open large language model for code with multi-turn program synthesis
Original
Erik Nijkamp, Bo Pang, Hiroaki Hayashi, Lifu Tu, Huan Wang, Yingbo Zhou, Silvio Savarese, and Caiming Xiong. 2022 · 2022
Later among the works it cites.
A Taxonomy of Prompt Modifiers for Text-To-Image Generation
Jonas Oppenlaender. 2022 · 2022
Later among the works it cites.
High-Resolution Image Synthesis with Latent Diffusion Models. In Proceedings of the Computer Vision and Pattern Recognition Conference . 10684–10695
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer. 2022 · 2022
Later among the works it cites.
Galactica: A large language model for science
Original
Ross Taylor, Marcin Kardas, Guillem Cucurull, Thomas Scialom, Anthony Hartshorn, Elvis Saravia, Andrew Poulton, Viktor Kerkez, and Robert Stojnic. 2022 · 2022
Later among the works it cites.
GSR: A Generalized Symbolic Regression Approach
Tony Tohme, Dehong Liu, and Kamal Youcef-Toumi. 2022 · 2022
Later among the works it cites.
Transformers learn in-context by gradient descent
Original
Johannes von Oswald, Eyvind Niklasson, Ettore Randazzo, João Sacramento, Alexander Mordvintsev, Andrey Zhmoginov, and Max Vladymyrov. 2022 · 2022
Later among the works it cites.
Evaluate & Evaluation on the Hub: Better Best Practices for Data and Model Measurement
Original
Leandro von Werra, Lewis Tunstall, Abhishek Thakur, Alexandra Sasha Luccioni, Tristan Thrush, Aleksandra Piktus, Felix Marty, Nazneen Rajani, Victor Mustar, Helen Ngo, et al · 2022
Later among the works it cites.
Emergent abilities of large language models
Original
Jason Wei, Yi Tay, Rishi Bommasani, Colin Raffel, Barret Zoph, Sebastian Borgeaud, Dani Yogatama, Maarten Bosma, Denny Zhou, Donald Metzler, et al · 2022
Later among the works it cites.
Using denoising autoencoder genetic programming to control exploration and exploitation in search. In European Conference on Genetic Programming (Part of EvoStar) . Springer, 102–117
David Wittenberg. 2022 · 2022
Later among the works it cites.
An Explanation of In-context Learning as Implicit Bayesian Inference. In International Conference on Learning Representations
Sang Michael Xie, Aditi Raghunathan, Percy Liang, and Tengyu Ma. 2022 · 2022
Later among the works it cites.
CoCa: Contrastive Captioners are Image-Text Foundation Models
Jiahui Yu, Zirui Wang, Vijay Vasudevan, Legg Yeung, Mojtaba Seyedhosseini, and Yonghui Wu. 2022 · 2022
Later among the works it cites.
Language Model Crossover: Variation through Few-Shot Prompting
Original
Anonymous (Same authors as present manuscript). 2023 · 2023
Closest in time.
Pythia: A suite for analyzing large language models across training and scaling. In International Conference on Machine Learning . PMLR, 2397–2430
Stella Biderman, Hailey Schoelkopf, Quentin Gregory Anthony, Herbie Bradley, Kyle O’Brien, Eric Hallahan, Mohammad Aflah Khan, Shivanshu Purohit, USVSN Sai Prashanth, Edward Raff, et al · 2023
Closest in time.
EvoPrompting: Language Models for Code-Level Neural Architecture Search
Original
Angelica Chen, David M Dohan, and David R So. 2023 · 2023
Closest in time.
Towards automated circuit discovery for mechanistic interpretability
Arthur Conmy, Augustine Mavor-Parker, Aengus Lynch, Stefan Heimersheim, and Adrià Garriga-Alonso. 2023 · 2023
Closest in time.
Promptbreeder: Self-referential self-improvement via prompt evolution
Original
Chrisantha Fernando, Dylan Banarse, Henryk Michalewski, Simon Osindero, and Tim Rocktäschel. 2023 · 2023
Closest in time.
Looped transformers as programmable computers
Original
Angeliki Giannou, Shashank Rajput, Jy-yong Sohn, Kangwook Lee, Jason D Lee, and Dimitris Papailiopoulos. 2023 · 2023
Closest in time.
Connecting Large Language Models with Evolutionary Algorithms Yields Powerful Prompt Optimizers
Original
Qingyan Guo, Rui Wang, Junliang Guo, Bei Li, Kaitao Song, Xu Tan, Guoqing Liu, Jiang Bian, and Yujiu Yang. 2023 · 2023
Closest in time.
Evolution through large models
Joel Lehman, Jonathan Gordon, Shawn Jain, Kamal Ndousse, Cathy Yeh, and Kenneth O Stanley. 2023 · 2023
Closest in time.
Large Language Model for Multi-objective Evolutionary Optimization
Original
Fei Liu, Xi Lin, Zhenkun Wang, Shunyu Yao, Xialiang Tong, Mingxuan Yuan, and Qingfu Zhang. 2023a · 2023
Closest in time.
Algorithm Evolution Using Large Language Model
Original
Fei Liu, Xialiang Tong, Mingxuan Yuan, and Qingfu Zhang. 2023b · 2023
Closest in time.
Fully Autonomous Programming with Large Language Models. In Proceedings of the Genetic and Evolutionary Computation Conference (GECCO) . ACM
Vadim Liventsev, Anastasiia Grishina, Aki Härmä, and Leon Moonen. 2023 · 2023
Closest in time.
Eureka: Human-Level Reward Design via Coding Large Language Models
Original
Yecheng Jason Ma, William Liang, Guanzhi Wang, De-An Huang, Osbert Bastani, Dinesh Jayaraman, Yuke Zhu, Linxi Fan, and Anima Anandkumar. 2023 · 2023
Closest in time.
LLMatic: Neural Architecture Search via Large Language Models and Quality Diversity Optimization
Original
Muhammad U. Nasir, Sam Earle, Julian Togelius, Steven James, and Christopher Cleghorn. 2023 · 2023
Closest in time.
Mathematical discoveries from program search with large language models
Bernardino Romera-Paredes, Mohammadamin Barekatain, Alexander Novikov, Matej Balog, M Pawan Kumar, Emilien Dupont, Francisco JR Ruiz, Jordan S Ellenberg, Pengming Wang, Omar Fawzi, et al · 2023
Closest in time.
LAION-Aesthetics
Christoph Schuhmann. 2022 · 2023
Closest in time.
Memory augmented large language models are computationally universal
Original
Dale Schuurmans. 2023 · 2023
Closest in time.
Efficient llm inference on cpus
Original
Haihao Shen, Hanwen Chang, Bo Dong, Yu Luo, and Hengyu Meng. 2023 · 2023
Closest in time.
Mariogpt: Open-ended text2level generation through large language models
Shyam Sudhakaran, Miguel González-Duque, Matthias Freiberger, Claire Glanois, Elias Najarro, and Sebastian Risi. 2023 · 2023
Closest in time.
WizardLM: Empowering Large Language Models to Follow Complex Instructions
Original
Can Xu, Qingfeng Sun, Kai Zheng, Xiubo Geng, Pu Zhao, Jiazhan Feng, Chongyang Tao, and Daxin Jiang. 2023 · 2023
Closest in time.
Large Language Models as Optimizers
Original
Chengrun Yang, Xuezhi Wang, Yifeng Lu, Hanxiao Liu, Quoc V. Le, Denny Zhou, and Xinyun Chen. 2023 · 2023
Closest in time.
Rethinking Interpretability in the Era of Large Language Models
Original
Chandan Singh, Jeevana Priya Inala, Michel Galley, Rich Caruana, and Jianfeng Gao. 2024 · 2024
Closest in time.
NoMAD-Attention: Efficient LLM Inference on CPUs Through Multiply-add-free Attention
Original
Tianyi Zhang, Jonah Wonkyu Yi, Bowen Yao, Zhaozhuo Xu, and Anshumali Shrivastava. 2024 · 2024
Closest in time.