Fetching the paper…
Reading the bibliography…
Solving symbolic mathematics has always been of in the arena of human ingenuity that needs compositional reasoning and recurrence.
Symbolic integration
Joel Moses · 1967
Earlier work this paper cites.
Guns, Germs, and Steel: the Fates of Human Societies
Jared M. Diamond · 1998
Earlier work this paper cites.
A neural probabilistic language model
Yoshua Bengio, Réjean Ducharme, Pascal Vincent, and Christian Janvin · 2003
Earlier work this paper cites.
Developmental Plasticity and Evolution
M.J. West-Eberhard and Oxford University Press · 2003
Earlier work this paper cites.
Critical phenomena in natural sciences: chaos, fractals, selforganization and disorder: concepts and tools
Didier Sornette · 2006
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2014
Earlier work this paper cites.
Learning to discover efficient mathematical identities, 2014
Wojciech Zaremba, Karol Kurach, and Rob Fergus · 2014
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2015
Earlier work this paper cites.
Attention-based models for speech recognition
Jan Chorowski, Dzmitry Bahdanau, Dmitriy Serdyuk, Kyunghyun Cho, and Yoshua Bengio · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
Rethinking the inception architecture for computer vision, 2015
Christian Szegedy, Vincent Vanhoucke, Sergey Ioffe, Jonathon Shlens, and Zbigniew Wojna · 2015
Earlier work this paper cites.
Learning to execute, 2015
Wojciech Zaremba and Ilya Sutskever · 2015
Earlier work this paper cites.
Retain: An interpretable predictive model for healthcare using reverse time attention mechanism
Edward Choi, Mohammad Taha Bahadori, Joshua A Kulas, Andy Schuetz, Walter F Stewart, and Jimeng Sun · 2016
Earlier work this paper cites.
Neural gpus learn algorithms, 2016
Łukasz Kaiser and Ilya Sutskever · 2016
Earlier work this paper cites.
Learning continuous semantic representations of symbolic expressions, 2017
Miltiadis Allamanis, Pankajan Chanthirasegaran, Pushmeet Kohli, and Charles Sutton · 2017
Earlier work this paper cites.
Program induction by rationale generation : Learning to solve and explain algebraic word problems, 2017
Wang Ling, Dani Yogatama, Chris Dyer, and Phil Blunsom · 2017
Earlier work this paper cites.
Deep network guided proof search, 2017
Sarah Loos, Geoffrey Irving, Christian Szegedy, and Cezary Kaliszyk · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Stress and stability: applying the anna karenina principle to animal microbiomes
Jesse Zaneveld, Ryan McMinds, and Rebecca Vega Thurber · 2017
Earlier work this paper cites.
Towards solving differential equations through neural programming
Forough Arabshahi, Sameer Singh, and Animashree Anandkumar · 2018
Cited alongside, same era.
The lottery ticket hypothesis: Training pruned neural networks
Jonathan Frankle and Michael Carbin · 2018
Cited alongside, same era.
Marian: Fast neural machine translation in C++
Marcin Junczys-Dowmunt, Roman Grundkiewicz, Tomasz Dwojak, Hieu Hoang, Kenneth Heafield, Tom Neckermann, Frank Seide, Ulrich Germann, Alham Fikri Aji, Nikolay Bogoychev, André F. T. Martins, and Alexandra Birch · 2018
Cited alongside, same era.
Visualizing the loss landscape of neural nets
Hao Li, Zheng Xu, Gavin Taylor, Christoph Studer, and Tom Goldstein · 2018
Cited alongside, same era.
Implicit self-regularization in deep neural networks: Evidence from random matrix theory and implications for learning, 2018
Charles H. Martin and Michael W. Mahoney · 2018
Cited alongside, same era.
Object detection based on an adaptive attention mechanism
Wei Li, Kai Liu, Lizhe Zhang, and Fei Cheng · 2020
Later among the works it cites.
Multilingual denoising pre-training for neural machine translation
Yinhan Liu, Jiatao Gu, Naman Goyal, Xian Li, Sergey Edunov, Marjan Ghazvininejad, Mike Lewis, and Luke Zettlemoyer · 2020
Later among the works it cites.
Generative language modeling for automated theorem proving, 2020
Stanislas Polu and Ilya Sutskever · 2020
Later among the works it cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu · 2020
Later among the works it cites.
Revisiting model stitching to compare neural representations
Yamini Bansal, Preetum Nakkiran, and Boaz Barak · 2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Neural arithmetic logic units, 2018
Andrew Trask, Felix Hill, Scott Reed, Jack Rae, Chris Dyer, and Phil Blunsom · 2018
Cited alongside, same era.
The use of deep learning for symbolic integration: A review of (lample and charton, 2019), 2019
Ernest Davis · 2019
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding, 2019
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Cited alongside, same era.
Attention branch network: Learning of attention mechanism for visual explanation
Hiroshi Fukui, Tsubasa Hirakawa, Takayoshi Yamashita, and Hironobu Fujiyoshi · 2019
Cited alongside, same era.
Deep learning for symbolic mathematics
Guillaume Lample and François Charton · 2019
Cited alongside, same era.
Analysing mathematical reasoning abilities of neural models, 2019
David Saxton, Edward Grefenstette, Felix Hill, and Pushmeet Kohli · 2019
Cited alongside, same era.
Videobert: A joint model for video and language representation learning
Chen Sun, Austin Myers, Carl Vondrick, Kevin Murphy, and Cordelia Schmid · 2019
Cited alongside, same era.
Linear algebra with transformers, 2021
François Charton · 2021
Closest in time.
Pre-trained image processing transformer
Hanting Chen, Yunhe Wang, Tianyu Guo, Chang Xu, Yiping Deng, Zhenhua Liu, Siwei Ma, Chunjing Xu, Chao Xu, and Wen Gao · 2021
Closest in time.
Advancing mathematics by guiding human intuition with ai
Alex Davies, Petar Veličković, Lars Buesing, Sam Blackwell, Daniel Zheng, Nenad Tomašev, Richard Tanburn, Peter Battaglia, Charles Blundell, András Juhász, Marc Lackenby, Geordie Williamson, Demis Hassabis, and Pushmeet Kohli · 2021
Closest in time.
Manoj Kumar, Dirk Weissenborn, and Nal Kalchbrenner · 2021
Closest in time.
Pretrained transformers as universal computation engines
Kevin Lu, Aditya Grover, Pieter Abbeel, and Igor Mordatch · 2021
Closest in time.
Symbolicgpt: A generative transformer model for symbolic regression, 2021
Mojtaba Valipour, Bowen You, Maysum Panju, and Ali Ghodsi · 2021
Closest in time.
A neural network solves, explains, and generates university math problems by program synthesis and few-shot learning at human level
Iddo Drori, Sarah Zhang, Reece Shuttleworth, Leonard Tang, Albert Lu, Elizabeth Ke, Kevin Liu, Linda Chen, Sunny Tran, Newman Cheng, Roman Wang, Nikhil Singh, Taylor L. Patti, Jayson Lynch, Avi Shporer, Nakul Verma, Eugene Wu, and Gilbert Strang · 2022
Closest in time.
Solving quantitative reasoning problems with language models
Aitor Lewkowycz, Anders Johan Andreassen, David Dohan, Ethan Dyer, Henryk Michalewski, Vinay Venkatesh Ramasesh, Ambrose Slone, Cem Anil, Imanol Schlag, Theo Gutman-Solo, Yuhuai Wu, Behnam Neyshabur, Guy Gur-Ari, and Vedant Misra · 2022
Closest in time.
Lemma: Bootstrapping high-level mathematical reasoning with learned symbolic abstractions, 2022
Zhening Li, Gabriel Poesia, Omar Costilla-Reyes, Noah Goodman, and Armando Solar-Lezama · 2022
Closest in time.
Emergent analogical reasoning in large language models, 2022
Taylor Webb, Keith J. Holyoak, and Hongjing Lu · 2022
Closest in time.
Emergent abilities of large language models
Jason Wei, Yi Tay, Rishi Bommasani, Colin Raffel, Barret Zoph, Sebastian Borgeaud, Dani Yogatama, Maarten Bosma, Denny Zhou, Donald Metzler, Ed H. Chi, Tatsunori Hashimoto, Oriol Vinyals, Percy Liang, Jeff Dean, and William Fedus · 2022
Closest in time.
Autoformalization with large language models
Yuhuai Wu, Albert Qiaochu Jiang, Wenda Li, Markus Norman Rabe, Charles E Staats, Mateja Jamnik, and Christian Szegedy · 2022
Closest in time.
Anna karenina principle
Wikipedia · 2023
Closest in time.