Fetching the paper…
Reading the bibliography…
Reasoning in mathematical domains remains a significant challenge for relatively small language models (LMs).
O znaczeniu i potrzebach logiki matematycznej
Jan Łukasiewicz · 1929
Earlier work this paper cites.
Analyzing the effectiveness and applicability of co-training
Kamal Nigam and Rayid Ghani · 2000
Earlier work this paper cites.
Multi-view clustering
Steffen Bickel and Tobias Scheffer · 2004
Earlier work this paper cites.
Robust co-training
Shiliang Sun and Feng Jin · 2011
Earlier work this paper cites.
Learning to solve arithmetic word problems with verb categorization
Mohammad Javad Hosseini, Hannaneh Hajishirzi, Oren Etzioni, and Nate Kushman · 2014
Earlier work this paper cites.
Learning to automatically solve algebra word problems
Nate Kushman, Luke Zettlemoyer, Regina Barzilay, and Yoav Artzi · 2014
Earlier work this paper cites.
Diversity-induced multi-view subspace clustering
Xiaochun Cao, Changqing Zhang, Huazhu Fu, Si Liu, and Hua Zhang · 2015
Earlier work this paper cites.
Parsing algebraic word problems into equations
Rik Koncel-Kedziorski, Hannaneh Hajishirzi, Ashish Sabharwal, Oren Etzioni, and Siena Dumas Ang · 2015
Earlier work this paper cites.
Large-scale multi-view spectral clustering via bipartite graph
Yeqing Li, Feiping Nie, Heng Huang, and Junzhou Huang · 2015
Earlier work this paper cites.
Solving general arithmetic word problems
Subhro Roy and Dan Roth · 2015
Earlier work this paper cites.
Mawps: A math word problem repository
Rik Koncel-Kedziorski, Subhro Roy, Aida Amini, Nate Kushman, and Hannaneh Hajishirzi · 2016
Earlier work this paper cites.
Translating a math word problem to a expression tree
Lei Wang, Yan Wang, Deng Cai, Dongxiang Zhang, and Xiaojiang Liu · 2018
Earlier work this paper cites.
Mathqa: Towards interpretable math word problem solving with operation-based formalisms
Aida Amini, Saadia Gabriel, Shanchuan Lin, Rik Koncel-Kedziorski, Yejin Choi, and Hannaneh Hajishirzi · 2019
Earlier work this paper cites.
A goal-driven tree-structured neural model for math word problems
Zhipeng Xie and Shichao Sun · 2019
Earlier work this paper cites.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Earlier work this paper cites.
A diverse corpus for evaluating and developing english math word problem solvers
Shen-Yun Miao, Chao-Chun Liang, and Keh-Yih Su · 2020
Cited alongside, same era.
Structural information preserving for graph-to-text generation
Linfeng Song, Ante Wang, Jinsong Su, Yue Zhang, Kun Xu, Yubin Ge, and Dong Yu · 2020
Cited alongside, same era.
Graph-to-tree learning for solving math word problems
Jipeng Zhang, Lei Wang, Roy Ka-Wei Lee, Yi Bin, Yan Wang, Jie Shao, and Ee-Peng Lim · 2020
Cited alongside, same era.
Ape210k: A large-scale and template-rich dataset of math word problems
Wei Zhao, Mingyue Shang, Yang Liu, Liang Wang, and Jingming Liu · 2020
Cited alongside, same era.
Training verifiers to solve math word problems
Karl Cobbe, Vineet Kosaraju, Mohammad Bavarian, Mark Chen, Heewoo Jun, Lukasz Kaiser, Matthias Plappert, Jerry Tworek, Jacob Hilton, Reiichiro Nakano, et al · 2021
Cited alongside, same era.
Mwp-bert: Numeracy-augmented pre-training for math word problem solving
Zhenwen Liang, Jipeng Zhang, Lei Wang, Wei Qin, Yunshi Lan, Jie Shao, and Xiangliang Zhang · 2022
Later among the works it cites.
Teaching small language models to reason
Lucie Charlotte Magister, Jonathan Mallinson, Jakub Adamek, Eric Malmi, and Aliaksei Severyn · 2022
Later among the works it cites.
Crosslingual generalization through multitask finetuning
Niklas Muennighoff, Thomas Wang, Lintang Sutawika, Adam Roberts, Stella Biderman, Teven Le Scao, M Saiful Bari, Sheng Shen, Zheng-Xin Yong, Hailey Schoelkopf, et al · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning by fixing: Solving math word problems with weak supervision
Yining Hong, Qing Li, Daniel Ciao, Siyuan Huang, and Song-Chun Zhu · 2021
Cited alongside, same era.
Show your work: Scratchpads for intermediate computation with language models
Maxwell Nye, Anders Johan Andreassen, Guy Gur-Ari, Henryk Michalewski, Jacob Austin, David Bieber, David Dohan, Aitor Lewkowycz, Maarten Bosma, David Luan, Charles Sutton, and Augustus Odena · 2021
Cited alongside, same era.
Are nlp models really able to solve simple math word problems?
Arkil Patel, Satwik Bhattamishra, and Navin Goyal · 2021
Cited alongside, same era.
Neural-symbolic solver for math word problems with auxiliary tasks
Jinghui Qin, Xiaodan Liang, Yining Hong, Jianheng Tang, and Liang Lin · 2021
Cited alongside, same era.
Self-teaching machines to read and comprehend with large-scale multi-subject question-answering data
Dian Yu, Kai Sun, Dong Yu, and Claire Cardie · 2021
Cited alongside, same era.
Large language models are reasoning teachers
Namgyu Ho, Laura Schmid, and Se-Young Yun · 2022
Cited alongside, same era.
Learning to reason deductively: Math word problem solving as complex relation extraction
Zhanming Jie, Jierui Li, and Wei Lu · 2022
Cited alongside, same era.
Kumar Shridhar, Alessandro Stolfo, and Mrinmaya Sachan · 2022
Later among the works it cites.
The ai teacher test: Measuring the pedagogical ability of blender and gpt-3 in educational dialogues
Anaïs Tack and Chris Piech · 2022
Later among the works it cites.
Investigating math word problems using pretrained multilingual language models
Minghuan Tan, Lei Wang, Lingxiao Jiang, and Jing Jiang · 2022
Later among the works it cites.
Multi-view self-attention based transformer for speaker recognition
Rui Wang, Junyi Ao, Long Zhou, Shujie Liu, Zhihua Wei, Tom Ko, Qing Li, and Yu Zhang · 2022
Later among the works it cites.
Towards understanding ensemble, knowledge distillation and self-distillation in deep learning
Zeyuan Allen-Zhu and Yuanzhi Li · 2023
Closest in time.
Specializing smaller language models towards multi-step reasoning
Yao Fu, Hao Peng, Litu Ou, Ashish Sabharwal, and Tushar Khot · 2023
Closest in time.
Cheng-Yu Hsieh, Chun-Liang Li, Chih-Kuan Yeh, Hootan Nakhost, Yasuhisa Fujii, Alexander Ratner, Ranjay Krishna, Chen-Yu Lee, and Tomas Pfister · 2023
Closest in time.
Let gpt be a math tutor: Teaching math word problem solvers with customized exercise generation
Zhenwen Liang, Wenhao Yu, Tanmay Rajpurohit, Peter Clark, Xiangliang Zhang, and Ashwin Kaylan · 2023
Closest in time.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al · 2023
Closest in time.
Knowledge graph augmented network towards multiview representation learning for aspect-based sentiment analysis
Qihuang Zhong, Liang Ding, Juhua Liu, Bo Du, Hua Jin, and Dacheng Tao · 2023
Closest in time.