Fetching the paper…
Reading the bibliography…
Pre-trained large language models (LLMs) exhibit powerful capabilities for generating natural text.
Goldberg, D.E.: Genetic Algorithms. pearson education India, New York (1989)
1989
Earlier work this paper cites.
Deb, K., Agrawal, R.B., et al
1995
Earlier work this paper cites.
Deb, K., Pratap, A., Agarwal, S., Meyarivan, T.: A fast and elitist multiobjective genetic algorithm: Nsga-ii. IEEE transactions on evolutionary computation 6
2002
Earlier work this paper cites.
Singh, G., Deb, K.: Comparison of multi-modal optimization algorithms based on evolutionary algorithms. In: Proceedings of the 8th Annual Conference on Genetic and Evolutionary Computation. GECCO ’06, pp. 1305–1312. Association for Computing Machinery, New York, NY, USA (2006)
2006
Earlier work this paper cites.
Turing, A.M.: Computing Machinery and Intelligence. Springer, Dordrecht (2009)
2009
Earlier work this paper cites.
Das, S., Suganthan, P.N.: Differential evolution: A survey of the state-of-the-art. IEEE transactions on evolutionary computation 15
2010
Earlier work this paper cites.
Das, S., Maity, S., Qu, B.-Y., Suganthan, P.N.: Real-parameter evolutionary multimodal optimization—a survey of the state-of-the-art. Swarm and Evolutionary Computation 1
2011
Earlier work this paper cites.
Simon, D.: Evolutionary Optimization Algorithms. John Wiley & Sons, Hoboken, New Jersey (2013)
2013
Earlier work this paper cites.
Burke, E.K., Gendreau, M., Hyde, M., Kendall, G., Ochoa, G., Özcan, E., Qu, R.: Hyper-heuristics: A survey of the state of the art. Journal of the Operational Research Society 64
2013
Earlier work this paper cites.
Wierstra, D., Schaul, T., Glasmachers, T., Sun, Y., Peters, J., Schmidhuber, J.: Natural evolution strategies. The Journal of Machine Learning Research 15
2014
Earlier work this paper cites.
Hirschberg, J., Manning, C.D.: Advances in natural language processing. Science 349
2015
Earlier work this paper cites.
Eiben, A.E., Smith, J.: From evolutionary computation to the evolution of things. Nature 521
2015
Earlier work this paper cites.
Gupta, A., Ong, Y.-S., Feng, L.: Multifactorial evolution: Toward evolutionary multitasking. IEEE Transactions on Evolutionary Computation 20
2016
Earlier work this paper cites.
Hansen, N.: The cma evolution strategy: A tutorial. arXiv preprint arXiv:1604.00772 (2016)
2016
Earlier work this paper cites.
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A.N., Kaiser, Ł., Polosukhin, I.: Attention is all you need. Advances in neural information processing systems 30
2017
Earlier work this paper cites.
Qian, H., Yu, Y.: Solving high-dimensional multi-objective optimization problems with low effective dimensions. In: Proceedings of the AAAI Conference on Artificial Intelligence, vol. 31 (2017)
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
Radford, A., Narasimhan, K., Salimans, T., Sutskever, I., et al.: Improving language understanding by generative pre-training. Preprint at https://openai.com/research/language-unsupervised (2018)
2018
Earlier work this paper cites.
Jin, Y., Wang, H., Chugh, T., Guo, D., Miettinen, K.: Data-driven evolutionary optimization: An overview and case studies. IEEE Transactions on Evolutionary Computation 23
2018
Earlier work this paper cites.
Gupta, A., Ong, Y.-S., Feng, L.: Insights on transfer optimization: Because experience is the best teacher. IEEE Transactions on Emerging Topics in Computational Intelligence 2
2018
Earlier work this paper cites.
Sener, O., Koltun, V.: Multi-task learning as multi-objective optimization. Advances in neural information processing systems 31
2018
Earlier work this paper cites.
Devlin, J., Chang, M.-W., Lee, K., Toutanova, K.: BERT: Pre-training of deep bidirectional transformers for language understanding. In: Burstein, J., Doran, C., Solorio, T. (eds.) Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers), pp. 4171–4186. Association for Computational Linguistics, Minneapolis, Minnesota (2019)
2019
Earlier work this paper cites.
Stanley, K.O., Clune, J., Lehman, J., Miikkulainen, R.: Designing neural networks through neuroevolution. Nature Machine Intelligence 1
2019
Earlier work this paper cites.
Voita, E., Talbot, D., Moiseev, F., Sennrich, R., Titov, I.: Analyzing multi-head self-attention: Specialized heads do the heavy lifting, the rest can be pruned. In: Korhonen, A., Traum, D., Màrquez, L. (eds.) Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics, pp. 5797–5808. Association for Computational Linguistics, Florence, Italy (2019)
2019
Earlier work this paper cites.
Correia, G.M., Niculae, V., Martins, A.F.T.: Adaptively sparse transformers. In: Inui, K., Jiang, J., Ng, V., Wan, X. (eds.) Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP), pp. 2174–2184. Association for Computational Linguistics, Hong Kong, China (2019)
2019
Earlier work this paper cites.
Zhou, Z.-H., Yu, Y., Qian, C.: Evolutionary Learning: Advances in Theories and Algorithms. Springer, Singapore (2019)
2019
Earlier work this paper cites.
Hassanat, A., Almohammadi, K., Alkafaween, E., Abunawas, E., Hammouri, A., Prasath, V.S.: Choosing mutation and crossover ratios for genetic algorithms—a review with a new dynamic approach. Information 10
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Dai, Z., Yang, Z., Yang, Y., Carbonell, J., Le, Q., Salakhutdinov, R.: Transformer-XL: Attentive language models beyond a fixed-length context. In: Korhonen, A., Traum, D., Màrquez, L. (eds.) Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics, pp. 2978–2988. Association for Computational Linguistics, Florence, Italy (2019)
2019
Earlier work this paper cites.
Lin, X., Zhen, H.-L., Li, Z., Zhang, Q.-F., Kwong, S.: Pareto multi-task learning. Advances in neural information processing systems 32
2019
Earlier work this paper cites.
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J.D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., Agarwal, S., Herbert-Voss, A., Krueger, G., Henighan, T., Child, R., Ramesh, A., Ziegler, D., Wu, J., Winter, C., Hesse, C., Chen, M., Sigler, E., Litwin, M., Gray, S., Chess, B., Clark, J., Berner, C., McCandlish, S., Radford, A., Sutskever, I., Amodei, D.: Language models are few-shot learners. Advances in neural information processing systems 33
2020
Earlier work this paper cites.
Bhojanapalli, S., Yun, C., Rawat, A.S., Reddi, S., Kumar, S.: Low-rank bottleneck in multi-head attention models. In: III, H.D., Singh, A. (eds.) Proceedings of the 37th International Conference on Machine Learning. Proceedings of Machine Learning Research, vol. 119, pp. 864–873 (2020)
2020
Earlier work this paper cites.
Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., Liu, P.J.: Exploring the limits of transfer learning with a unified text-to-text transformer. The Journal of Machine Learning Research 21
2020
Earlier work this paper cites.
He, P., Liu, X., Gao, J., Chen, W.: Deberta: Decoding-enhanced bert with disentangled attention. In: International Conference on Learning Representations (2020)
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
Li, J., Shang, S., Chen, L.: Domain generalization for named entity boundary detection via metalearning. IEEE Transactions on Neural Networks and Learning Systems 32
2021
Earlier work this paper cites.
Jin, Y., Wang, H., Sun, C.: Data-driven Evolutionary Optimization. Springer, Cham, Switzerland (2021)
2021
Earlier work this paper cites.
Miikkulainen, R., Forrest, S.: A biological perspective on evolutionary computation. Nature Machine Intelligence 3
2021
Earlier work this paper cites.
Zhang, J., Xu, C., Li, J., Chen, W., Wang, Y., Tai, Y., Chen, S., Wang, C., Huang, F., Liu, Y.: Analogous to evolutionary algorithm: Designing a unified sequence model. Advances in Neural Information Processing Systems 34
2021
Cited alongside, same era.
Tian, Y., Si, L., Zhang, X., Cheng, R., He, C., Tan, K.C., Jin, Y.: Evolutionary large-scale multi-objective optimization: A survey. ACM Computing Surveys (CSUR) 54
2021
Cited alongside, same era.
Niu, Z., Zhong, G., Yu, H.: A review on the attention mechanism of deep learning. Neurocomputing 452
2021
Cited alongside, same era.
Dong, Y., Cordonnier, J.-B., Loukas, A.: Attention is not all you need: Pure attention loses rank doubly exponentially with depth. In: International Conference on Machine Learning, pp. 2793–2803 (2021). PMLR
2021
Cited alongside, same era.
Chen, A., Dohan, D., So, D.: Evoprompting: Language models for code-level neural architecture search. In: Thirty-seventh Conference on Neural Information Processing Systems (2023)
2023
Later among the works it cites.
Zhang, Z., Wang, S., Yu, W., Xu, Y., Iter, D., Zeng, Q., Liu, Y., Zhu, C., Jiang, M.: Auto-instruct: Automatic instruction generation and ranking for black-box language models. In: Bouamor, H., Pino, J., Bali, K. (eds.) Findings of the Association for Computational Linguistics: EMNLP 2023, pp. 9850–9867. Association for Computational Linguistics, Singapore (2023)
2023
Later among the works it cites.
Fei, Z., Fan, M., Huang, J.: Gradient-free textual inversion. In: Proceedings of the 31st ACM International Conference on Multimedia. MM ’23, pp. 1364–1373. Association for Computing Machinery, New York, NY, USA (2023)
2023
Later among the works it cites.
Shen, M., Ghosh, S., Sattigeri, P., Das, S., Bu, Y., Wornell, G.: Reliable gradient-free and likelihood-free prompt tuning. In: Vlachos, A., Augenstein, I. (eds.) Findings of the Association for Computational Linguistics: EACL 2023, pp. 2416–2429. Association for Computational Linguistics, Dubrovnik, Croatia (2023)
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2021
Cited alongside, same era.
Li, J., Chiu, B., Shang, S., Shao, L.: Neural text segmentation and its application to sentiment analysis. IEEE Transactions on Knowledge and Data Engineering 34
2022
Cited alongside, same era.
Schramowski, P., Turan, C., Andersen, N., Rothkopf, C.A., Kersting, K.: Large pre-trained language models contain human-like biases of what is right and wrong to do. Nature Machine Intelligence 4
2022
Cited alongside, same era.
Sun, T., Shao, Y., Qian, H., Huang, X., Qiu, X.: Black-box tuning for language-model-as-a-service. In: International Conference on Machine Learning, pp. 20841–20855 (2022). PMLR
2022
Cited alongside, same era.
Kudela, J.: A critical problem in benchmarking and analysis of evolutionary computation methods. Nature Machine Intelligence 4
2022
Cited alongside, same era.
Wang, C., Liu, J., Wu, K., Wu, Z.: Solving multitask optimization problems with adaptive knowledge transfer via anomaly detection. IEEE Transactions on Evolutionary Computation 26
2022
Cited alongside, same era.
Lyu, S., Wu, X., Li, J., Chen, Q., Chen, H.: Do models learn the directionality of relations? a new evaluation: Relation direction recognition. IEEE Transactions on Emerging Topics in Computational Intelligence 6
2022
Cited alongside, same era.
Sun, T., He, Z., Qian, H., Zhou, Y., Huang, X.-J., Qiu, X.: Bbtv2: towards a gradient-free future with large language models. In: Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing, pp. 3916–3930 (2022)
2022
Cited alongside, same era.
2023
Later among the works it cites.
Qi, S., Zhang, Y.: Prompt-calibrated tuning: Improving black-box optimization for few-shot scenarios. In: 2023 4th International Seminar on Artificial Intelligence, Networking and Information Technology (AINIT), pp. 402–407 (2023). IEEE
2023
Later among the works it cites.
Han, C., Cui, L., Zhu, R., Wang, J., Chen, N., Sun, Q., Li, X., Gao, M.: When gradient descent meets derivative-free optimization: A match made in black-box scenario. In: Rogers, A., Boyd-Graber, J., Okazaki, N. (eds.) Findings of the Association for Computational Linguistics: ACL 2023, pp. 868–880. Association for Computational Linguistics, Toronto, Canada (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Prasad, A., Hase, P., Zhou, X., Bansal, M.: GrIPS: Gradient-free, edit-based instruction search for prompting large language models. In: Vlachos, A., Augenstein, I. (eds.) Proceedings of the 17th Conference of the European Chapter of the Association for Computational Linguistics, pp. 3845–3864. Association for Computational Linguistics, Dubrovnik, Croatia (2023)
2023
Later among the works it cites.
Zhou, H., Wan, X., Vulić, I., Korhonen, A.: Survival of the most influential prompts: Efficient black-box prompt search via clustering and pruning. In: Bouamor, H., Pino, J., Bali, K. (eds.) Findings of the Association for Computational Linguistics: EMNLP 2023, pp. 13064–13077. Association for Computational Linguistics, Singapore (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Yu, L., Chen, Q., Lin, J., He, L.: Black-box prompt tuning for vision-language model as a service. In: Proceedings of the Thirty-Second International Joint Conference on Artificial Intelligence, pp. 1686–1694 (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Zhang, M., Desai, N., Bae, J., Lorraine, J., Ba, J.: Using large language models for hyperparameter optimization. In: NeurIPS 2023 Foundation Models for Decision Making Workshop (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Lanzi, P.L., Loiacono, D.: Chatgpt and other large language models as evolutionary engines for online interactive collaborative game design. In: Proceedings of the Genetic and Evolutionary Computation Conference. GECCO ’23, pp. 1383–1390. Association for Computing Machinery, New York, NY, USA (2023)
2023
Later among the works it cites.
Sudhakaran, S., González-Duque, M., Freiberger, M., Glanois, C., Najarro, E., Risi, S.: MarioGPT: Open-ended text2level generation through large language models. In: Thirty-seventh Conference on Neural Information Processing Systems (2023)
2023
Later among the works it cites.
Jablonka, K.M., Ai, Q., Al-Feghali, A., Badhwar, S., Bocarsly, J.D., Bran, A.M., Bringuier, S., Brinson, L.C., Choudhary, K., Circi, D., et al
2023
Later among the works it cites.
2023
Later among the works it cites.
Li, J., Feng, S., Chiu, B.: Few-shot relation extraction with dual graph neural network interaction. IEEE Transactions on Neural Networks and Learning Systems 35
2024
Closest in time.
Jiao, L., Zhao, J., Wang, C., Liu, X., Liu, F., Li, L., Shang, R., Li, Y., Ma, W., Yang, S.: Nature-inspired intelligent computing: A comprehensive survey. Research 7
2024
Closest in time.
Lehman, J., Gordon, J., Jain, S., Ndousse, K., Yeh, C., Stanley, K.O.: In: Banzhaf, W., Machado, P., Zhang, M. (eds.) Evolution Through Large Models, pp. 331–366. Springer, Singapore (2024)
2024
Closest in time.
Liu, J., Sarker, R., Elsayed, S., Essam, D., Siswanto, N.: Large-scale evolutionary optimization: A review and comparative study. Swarm and Evolutionary Computation, 101466 (2024)
2024
Closest in time.
Su, J., Ahmed, M., Lu, Y., Pan, S., Bo, W., Liu, Y.: Roformer: Enhanced transformer with rotary position embedding. Neurocomputing 568
2024
Closest in time.
Ding, L., Zhang, J., Clune, J., Spector, L., Lehman, J.: Quality diversity through human feedback: Towards open-ended diversity-driven optimization. In: Forty-first International Conference on Machine Learning (2024)
2024
Closest in time.
2024
Closest in time.
Sun, Q., Han, C., Chen, N., Zhu, R., Gong, J., Li, X., Gao, M.: Make prompt-based black-box tuning colorful: Boosting model generalization from three orthogonal perspectives. In: Calzolari, N., Kan, M.-Y., Hoste, V., Lenci, A., Sakti, S., Xue, N. (eds.) Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024), pp. 10958–10969. ELRA and ICCL, Torino, Italia (2024)
2024
Closest in time.
Zheng, Y., Tan, Z., Li, P., Liu, Y.: Black-box prompt tuning with subspace learning. IEEE/ACM Transactions on Audio, Speech, and Language Processing 32
2024
Closest in time.
Ma, Y.J., Liang, W., Wang, G., Huang, D.-A., Bastani, O., Jayaraman, D., Zhu, Y., Fan, L., Anandkumar, A.: Eureka: Human-level reward design via coding large language models. In: The Twelfth International Conference on Learning Representations (2024)
2024
Closest in time.
Yang, C., Wang, X., Lu, Y., Liu, H., Le, Q.V., Zhou, D., Chen, X.: Large language models as optimizers. In: The Twelfth International Conference on Learning Representations (2024)
2024
Closest in time.
2024
Closest in time.
Xia, C.S., Paltenghi, M., Tian, J.L., Pradel, M., Zhang, L.: Fuzz4all: Universal fuzzing with large language models. In: 2024 IEEE/ACM 46th International Conference on Software Engineering (ICSE) (2024)
2024
Closest in time.