Language models as knowledge bases?
F. Petroni, T. Rocktäschel, P. Lewis, A. Bakhtin, Y. Wu, A. H. Miller, and S. Riedel · 2020
Later among the works it cites.
Generative language modeling for automated theorem proving, 2020
S. Polu and I. Sutskever · 2020
Later among the works it cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu · 2020
Later among the works it cites.
Leveraging pre-trained checkpoints for sequence generation tasks
S. Rothe, S. Narayan, and A. Severyn · 2020
Later among the works it cites.
A Promising Path Towards Autoformalization and General Artificial Intelligence , 2020
C. Szegedy, editor · 2020
Later among the works it cites.
Proofwriter: Generating implications, proofs, and abductive statements over natural language
Original
O. Tafjord, B. D. Mishra, and P. Clark · 2020
Later among the works it cites.
Learning to Prove Theorems by Learning to Generate Theorems
M. Wang and J. Deng · 2020
Later among the works it cites.
Exploration of neural machine translation in autoformalization of mathematics in Mizar
Q. Wang, C. Brown, C. Kaliszyk, and J. Urban · 2020
Later among the works it cites.
Transformers: State-of-the-art natural language processing
T. Wolf, L. Debut, V. Sanh, J. Chaumond, C. Delangue, A. Moi, P. Cistac, T. Rault, R. Louf, M. Funtowicz, J. Davison, S. Shleifer, P. von Platen, C. Ma, Y. Jernite, J. Plu, C. Xu, T. L. Scao, S. Gugger, M. Drame, Q. Lhoest, and A. M. Rush · 2020
Later among the works it cites.
Americasnli: Evaluating zero-shot natural language understanding of pretrained multilingual models in truly low-resource languages, 2021
A. Ebrahimi, M. Mager, A. Oncevay, V. Chaudhary, L. Chiruzzo, A. Fan, J. Ortega, R. Ramos, A. Rios, I. Vladimir, G. A. Giménez-Lugo, E. Mager, G. Neubig, A. Palmer, R. A. C. Solano, N. T. Vu, and K. Kann · 2021
Closest in time.
Measuring mathematical problem solving with the math dataset, 2021
D. Hendrycks, C. Burns, S. Kadavath, A. Arora, S. Basart, E. Tang, D. Song, and J. Steinhardt · 2021
Closest in time.
Isarstep: a benchmark for high-level mathematical reasoning
W. Li, L. Yu, Y. Wu, and L. C. Paulson · 2021
Closest in time.
Mathematical reasoning via self-supervised skip-tree training
M. N. Rabe, D. Lee, K. Bansal, and C. Szegedy · 2021
Closest in time.
Beir: A heterogenous benchmark for zero-shot evaluation of information retrieval models
Original
N. Thakur, N. Reimers, A. Rücklé, A. Srivastava, and I. Gurevych · 2021
Closest in time.
Lime: Learning inductive bias for primitives of mathematical reasoning, 2021
Y. Wu, M. Rabe, W. Li, J. Ba, R. Grosse, and C. Szegedy · 2021
Closest in time.