LSTM networks can perform dynamic counting
Mirac Suzgun, Yonatan Belinkov, Stuart Shieber, and Sebastian Gehrmann. 2019 · 2019
Later among the works it cites.
Quarel: A dataset and models for answering questions about qualitative relationships
Oyvind Tafjord, Peter Clark, Matt Gardner, Wen-tau Yih, and Ashish Sabharwal. 2019 · 2019
Later among the works it cites.
Do NLP models know numbers? probing numeracy in embeddings
Eric Wallace, Yizhong Wang, Sujian Li, Sameer Singh, and Matt Gardner. 2019 · 2019
Later among the works it cites.
Superglue: A stickier benchmark for general-purpose language understanding systems
Alex Wang, Yada Pruksachatkun, Nikita Nangia, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel Bowman. 2019 · 2019
Later among the works it cites.
An empirical investigation of contextualized number prediction
Taylor Berg-Kirkpatrick and Daniel Spokoyny. 2020 · 2020
Later among the works it cites.
On the computational power of transformers and its implications in sequence modeling
Satwik Bhattamishra, Arkil Patel, and Navin Goyal. 2020 · 2020
Later among the works it cites.
Language models are few-shot learners
Original
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2020
Later among the works it cites.
Injecting numerical reasoning skills into language models
Mor Geva, Ankit Gupta, and Jonathan Berant. 2020 · 2020
Later among the works it cites.
Learning numeral embedding
Chengyue Jiang, Zhonglin Nian, Kaihao Guo, Shanbo Chu, Yinggong Zhao, Libin Shen, and Kewei Tu. 2020 · 2020
Later among the works it cites.
Probing for multilingual numerical understanding in transformer-based language models
Devin Johnson, Denise Mak, Andrew Barker, and Lexi Loessberg-Zahl. 2020 · 2020
Later among the works it cites.
Birds have four legs?! NumerSense: Probing Numerical Commonsense Knowledge of Pre-Trained Language Models
Bill Yuchen Lin, Seyeon Lee, Rahul Khanna, and Xiang Ren. 2020 · 2020
Later among the works it cites.
Contextual datetime language model adaptation for speech recognition
Richard Diehl Martinez, Scott Novotney, Ivan Bulyko, Ariya Rastrow, and Andreas Stolcke. 2020 · 2020
Later among the works it cites.
Towards question format independent numerical reasoning: A set of prerequisite tasks
Original
Swaroop Mishra, Arindam Mitra, Neeraj Varshney, Bhavdeep Sachdeva, and Chitta Baral. 2020 · 2020
Later among the works it cites.
Dataset for evaluation of mathematical reasoning abilities in russian
Mikhail Nefedov. 2020 · 2020
Later among the works it cites.
Methods for numeracy-preserving word embeddings
Dhanasekar Sundararaman, Shijing Si, Vivek Subramanian, Guoyin Wang, Devamanyu Hazarika, and Lawrence Carin. 2020 · 2020
Later among the works it cites.
Leap-of-thought: Teaching pre-trained models to systematically reason over implicit knowledge
Alon Talmor, Oyvind Tafjord, Peter Clark, Yoav Goldberg, and Jonathan Berant. 2020 · 2020
Later among the works it cites.
Visual sense of number vs. sense of magnitude in humans and machines
Alberto Testolin, Serena Dolfi, Mathijs Rochus, and Marco Zorzi. 2020 · 2020
Later among the works it cites.
Do language embeddings capture scales?
Xikun Zhang, Deepak Ramachandran, Ian Tenney, Yanai Elazar, and Dan Roth. 2020 · 2020
Later among the works it cites.
Temporal common sense acquisition with minimal supervision
Ben Zhou, Qiang Ning, Daniel Khashabi, and Dan Roth. 2020 · 2020
Later among the works it cites.
Measuring mathematical problem solving with the math dataset
Original
Dan Hendrycks, Collin Burns, Saurav Kadavath, Akul Arora, Steven Basart, Eric Tang, Dawn Song, and Jacob Steinhardt. 2021 · 2021
Closest in time.
Investigating the limitations of the transformers with simple arithmetic tasks
Original
Rodrigo Nogueira, Zhiying Jiang, and Jimmy Li. 2021 · 2021
Closest in time.
Are nlp models really able to solve simple math word problems?
Original
Arkil Patel, Satwik Bhattamishra, and Navin Goyal. 2021 · 2021
Closest in time.