Fetching the paper…
Reading the bibliography…
We introduce SpreadsheetBench, a challenging spreadsheet manipulation benchmark exclusively derived from real-world scenarios, designed to immerse current large language models (LLMs) in the actual workflow of spreadsheet users.
John Wiley & Sons, 2006
M. Jackson and M. Staunton, Advanced modelling in finance using Excel and VBA · 2006
Earlier work this paper cites.
G. Cormack, I. Munro, T. Vasiga, and G. Kemkes, “Structure, scoring and purpose of computing competitions,” Informatics in education
2006
Earlier work this paper cites.
S. Gulwani, “Automating string processing in spreadsheets using input-output examples,” in Proceedings of the 38th annual ACM SIGPLAN-SIGACT symposium on Principles of programming languages
2011
Earlier work this paper cites.
S. Gulwani, “Automating string processing in spreadsheets using input-output examples,” ACM Sigplan Notices
2011
Earlier work this paper cites.
S. Gulwani, W. R. Harris, and R. Singh, “Spreadsheet data manipulation using examples,” Communications of the ACM
2012
Earlier work this paper cites.
John Wiley & Sons, 2014
W. L. Winston, Marketing analytics: Data-driven techniques with Microsoft Excel · 2014
Earlier work this paper cites.
P. Pasupat and P. Liang, “Compositional semantic parsing on semi-structured tables,” in Proceedings of the 53rd Annual Meeting of the Association for Computational Linguistics and the 7th International Joint Conference on Natural Language Processing (Volume 1: Long Papers)
2015
Earlier work this paper cites.
S.-C. Cheung, W. Chen, Y. Liu, and C. Xu, “Custodes: automatic spreadsheet cell clustering and smell detection using strong and weak features,” in Proceedings of the 38th International Conference on Software Engineering
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
T. Yu, R. Zhang, K. Yang, M. Yasunaga, D. Wang, Z. Li, J. Ma, I. Li, Q. Yao, S. Roman, et al
2018
Earlier work this paper cites.
S. Wasik, M. Antczak, J. Badura, A. Laskowski, and T. Sternal, “A survey on online judge systems and their applications,” ACM Computing Surveys (CSUR)
2018
Earlier work this paper cites.
S. Rahman, K. Mack, M. Bendre, R. Zhang, K. Karahalios, and A. Parameswaran, “Benchmarking spreadsheet systems,” in Proceedings of the 2020 acm sigmod international conference on management of data
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
J. Herzig, P. K. Nowak, T. Mueller, F. Piccinno, and J. Eisenschlos, “Tapas: Weakly supervised table parsing via pre-training,” in Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics
2020
Earlier work this paper cites.
X. Chen, P. Maniatis, R. Singh, C. Sutton, H. Dai, M. Lin, and D. Zhou, “Spreadsheetcoder: Formula prediction from semi-structured context,” in International Conference on Machine Learning
2021
Cited alongside, same era.
A. Elangovan, J. He, and K. Verspoor, “Memorization vs. generalization: Quantifying data leakage in nlp performance evaluation,” in Proceedings of the 16th Conference of the European Chapter of the Association for Computational Linguistics: Main Volume
2021
Cited alongside, same era.
F. Zhu, W. Lei, Y. Huang, C. Wang, S. Zhang, J. Lv, F. Feng, and T.-S. Chua, “Tat-qa: A question answering benchmark on a hybrid of tabular and textual content in finance,” in Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers)
2021
Cited alongside, same era.
P. Vaithilingam, T. Zhang, and E. L. Glassman, “Expectation vs. experience: Evaluating the usability of code generation tools powered by large language models,” in Chi conference on human factors in computing systems extended abstracts
Y. Zhao, H. Zhang, S. Si, L. Nan, X. Tang, and A. Cohan, “Investigating table-to-text generation capabilities of large language models in real-world information seeking scenarios,” in Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing: Industry Track
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2022
Cited alongside, same era.
L. Nan, C. Hsieh, Z. Mao, X. V. Lin, N. Verma, R. Zhang, W. Kryściński, H. Schoelkopf, R. Kong, X. Tang, et al
2022
Cited alongside, same era.
Y. Katsis, S. Chemmengath, V. Kumar, S. Bharadwaj, M. Canim, M. Glass, A. Gliozzo, F. Pan, J. Sen, K. Sankaranarayanan, et al
2022
Cited alongside, same era.
Z. Cheng, H. Dong, Z. Wang, R. Jia, J. Guo, Y. Gao, S. Han, J.-G. Lou, and D. Zhang, “Hitab: A hierarchical table dataset for question answering and natural language generation,” in Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
2022
Cited alongside, same era.
Q. Liu, B. Chen, J. Guo, M. Ziyadi, Z. Lin, W. Chen, and J.-G. Lou, “TAPEX: Table pre-training via learning a neural SQL executor,” in International Conference on Learning Representations
2022
Cited alongside, same era.
S. Yao, J. Zhao, D. Yu, N. Du, I. Shafran, K. R. Narasimhan, and Y. Cao, “React: Synergizing reasoning and acting in language models,” in The Eleventh International Conference on Learning Representations
2022
Cited alongside, same era.
2023
Cited alongside, same era.
J. Payan, S. Mishra, M. Singh, C. Negreanu, C. Poelitz, C. Baral, S. Roy, R. Chakravarthy, B. Van Durme, and E. Nouri, “Instructexcel: A benchmark for natural language instruction in excel,” in Findings of the Association for Computational Linguistics: EMNLP 2023
2023
Cited alongside, same era.
Y. Wang, Y. Kordi, S. Mishra, A. Liu, N. A. Smith, D. Khashabi, and H. Hajishirzi, “Self-instruct: Aligning language models with self-generated instructions,” in The 61st Annual Meeting Of The Association For Computational Linguistics
2023
Cited alongside, same era.
2024
Closest in time.
H. Li, J. Su, Y. Chen, Q. Li, and Z.-X. ZHANG, “Sheetcopilot: Bringing software productivity to the next level through large language models,” Advances in Neural Information Processing Systems
2024
Closest in time.
2024
Closest in time.
X. He, M. Zhou, X. Xu, X. Ma, R. Ding, L. Du, Y. Gao, R. Jia, X. Chen, S. Han, et al
2024
Closest in time.
2024
Closest in time.
X. Hu, Z. Zhao, S. Wei, Z. Chai, G. Wang, X. Wang, J. Su, J. Xu, M. Zhu, Y. Cheng, et al
2024
Closest in time.
2024
Closest in time.
D. Guo, Q. Zhu, D. Yang, Z. Xie, K. Dong, W. Zhang, G. Chen, X. Bi, Y. Wu, Y. Li, et al
2024
Closest in time.
2024
Closest in time.
AI@Meta, “Llama 3 model card,” 2024
2024
Closest in time.