Fetching the paper…
Reading the bibliography…
Pre-training has been investigated to improve the efficiency and performance of training neural operators in data-scarce settings.
Introduction to partial differential equations with applications
Zachmanoglou, E. C. and Thoe, D. W · 1986
Earlier work this paper cites.
Scheduled sampling for sequence prediction with recurrent neural networks
Bengio, S., Vinyals, O., Jaitly, N., and Shazeer, N · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Ronneberger, O., Fischer, P., and Brox, T · 2015
Earlier work this paper cites.
Improving language understanding by generative pre-training
Radford, A., Narasimhan, K., Salimans, T., Sutskever, I., et al · 2018
Earlier work this paper cites.
Group normalization
Wu, Y. and He, K · 2018
Earlier work this paper cites.
Variational physics-informed neural networks for solving partial differential equations
Kharazmi, E., Zhang, Z., and Karniadakis, G. E · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Radford, A., Wu, J., Child, R., Luan, D., Amodei, D., Sutskever, I., et al · 2019
Earlier work this paper cites.
Physics-informed neural networks: A deep learning framework for solving forward and inverse problems involving nonlinear partial differential equations
Raissi, M., Perdikaris, P., and Karniadakis, G. E · 2019
Earlier work this paper cites.
Bridging the gap between training and inference for neural machine translation
Zhang, W., Feng, Y., Meng, F., You, D., and Liu, Q · 2019
Earlier work this paper cites.
Language models are few-shot learners
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al · 2020
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Dosovitskiy, A., Beyer, L., Kolesnikov, A., Weissenborn, D., Zhai, X., Unterthiner, T., Dehghani, M., Minderer, M., Heigold, G., Gelly, S., et al · 2020
Earlier work this paper cites.
Momentum contrast for unsupervised visual representation learning
He, K., Fan, H., Wu, Y., Xie, S., and Girshick, R · 2020
Earlier work this paper cites.
Fourier neural operator for parametric partial differential equations
Li, Z., Kovachki, N., Azizzadenesheli, K., Liu, B., Bhattacharya, K., Stuart, A., and Anandkumar, A · 2020
Earlier work this paper cites.
Choose a transformer: Fourier or galerkin
Cao, S · 2021
Earlier work this paper cites.
Adaptive fourier neural operators: Efficient token mixers for transformers
Guibas, J., Mardani, M., Li, Z., Tao, A., Anandkumar, A., and Catanzaro, B · 2021
Earlier work this paper cites.
Highly accurate protein structure prediction with alphafold
Jumper, J., Evans, R., Pritzel, A., Green, T., Figurnov, M., Ronneberger, O., Tunyasuvunakool, K., Bates, R., Žídek, A., Potapenko, A., et al · 2021
Earlier work this paper cites.
Physics-informed machine learning
Karniadakis, G. E., Kevrekidis, I. G., Lu, L., Perdikaris, P., Wang, S., and Yang, L · 2021
Cited alongside, same era.
Physics-informed neural operator for learning partial differential equations
Li, Z., Zheng, H., Kovachki, N., Jin, D., Chen, H., Liu, B., Azizzadenesheli, K., and Anandkumar, A · 2021
Cited alongside, same era.
Learning nonlinear operators via deeponet based on the universal approximation theorem of operators
Lu, L., Jin, P., Pang, G., Zhang, Z., and Karniadakis, G. E · 2021
Cited alongside, same era.
Mlp-mixer: An all-mlp architecture for vision
Tolstikhin, I. O., Houlsby, N., Kolesnikov, A., Beyer, L., Zhai, X., Unterthiner, T., Yung, J., Steiner, A., Keysers, D., Uszkoreit, J., et al · 2021
Cited alongside, same era.
Factorized fourier neural operators
Tran, A., Mathews, A., Xie, L., and Ong, C. S · 2021
Cited alongside, same era.
Nuno: A general framework for learning parametric pdes with non-uniform data
Liu, S., Hao, Z., Ying, C., Su, H., Cheng, Z., and Zhu, J · 2023
Later among the works it cites.
Cfdbench: A comprehensive benchmark for machine learning methods in fluid dynamics
Luo, Y., Chen, Y., and Zhang, Z · 2023
Later among the works it cites.
Multiple physics pretraining for physical surrogate models
McCabe, M., Blancard, B. R.-S., Parker, L. H., Ohana, R., Cranmer, M., Bietti, A., Eickenberg, M., Golkar, S., Krawezik, G., Lanusse, F., et al · 2023
Later among the works it cites.
Self-supervised learning with lie symmetries for partial differential equations
Mialon, G., Garrido, Q., Lawrence, H., Rehman, D., LeCun, Y., and Kiani, B · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Message passing neural pde solvers
Brandstetter, J., Worrall, D., and Welling, M · 2022
Cited alongside, same era.
Scientific machine learning through physics–informed neural networks: Where we are and what’s next
Cuomo, S., Di Cola, V. S., Giampaolo, F., Rozza, G., Raissi, M., and Piccialli, F · 2022
Cited alongside, same era.
Towards multi-spatiotemporal-scale generalized pde modeling
Gupta, J. K. and Brandstetter, J · 2022
Cited alongside, same era.
Masked autoencoders are scalable vision learners
He, K., Chen, X., Xie, S., Li, Y., Dollár, P., and Girshick, R · 2022
Cited alongside, same era.
Pathak, J., Subramanian, S., Harrington, P., Raja, S., Chattopadhyay, A., Mardani, M., Kurth, T., Hall, D., Li, Z., Azizzadenesheli, K., et al · 2022
Cited alongside, same era.
Pdebench: An extensive benchmark for scientific machine learning
Takamoto, M., Praditia, T., Leiteritz, R., MacKinlay, D., Alesiani, F., Pflüger, D., and Niepert, M · 2022
Cited alongside, same era.
Sumformer: Universal approximation for efficient transformers
Alberti, S., Dern, N., Thesing, L., and Kutyniok, G · 2023
Cited alongside, same era.
Michałowska, K., Goswami, S., Karniadakis, G. E., and Riemer-Sørensen, S · 2023
Later among the works it cites.
Ai foundation models for weather and climate: Applications, design, and implementation
Mukkavilli, S. K., Civitarese, D. S., Schmude, J., Jakubik, J., Jones, A., Nguyen, N., Phillips, C., Roy, S., Singh, S., Watson, C., et al · 2023
Later among the works it cites.
Climax: A foundation model for weather and climate
Nguyen, T., Brandstetter, J., Kapoor, A., Gupta, J. K., and Grover, A · 2023
Later among the works it cites.
Vito: Vision transformer-operator
Ovadia, O., Kahana, A., Stinis, P., Turkel, E., and Karniadakis, G. E · 2023
Later among the works it cites.
Convolutional neural operators
Raonić, B., Molinaro, R., Rohner, T., Mishra, S., and de Bezenac, E · 2023
Later among the works it cites.
Subramanian, S., Harrington, P., Keutzer, K., Bhimji, W., Morozov, D., Mahoney, M., and Gholami, A · 2023
Later among the works it cites.
Solving high-dimensional pdes with latent spectral models
Wu, H., Hu, T., Luo, H., Wang, J., and Long, M · 2023
Later among the works it cites.
In-context operator learning with data prompts for differential equation problems
Yang, L., Liu, S., Meng, T., and Osher, S. J · 2023
Later among the works it cites.
Uni-mol: a universal 3d molecular representation learning framework
Zhou, G., Gao, Z., Ding, Q., Zheng, H., Xu, H., Wei, Z., Zhang, L., and Ke, G · 2023
Later among the works it cites.
Deep neural operators as accurate surrogates for shape optimization
Shukla, K., Oommen, V., Peyvan, A., Penwarden, M., Plewacki, N., Bravo, L., Ghoshal, A., Kirby, R. M., and Karniadakis, G. E · 2024
Closest in time.
Recfno: a resolution-invariant flow and heat field reconstruction method from sparse observations via fourier neural operator
Zhao, X., Chen, X., Gong, Z., Zhou, W., Yao, W., and Zhang, Y · 2024
Closest in time.