Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have demonstrated impressive capabilities across a range of scientific tasks including mathematics, physics, and chemistry.
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu, “Bleu: a method for automatic evaluation of machine translation,” in Proceedings of the 40th annual meeting of the Association for Computational Linguistics , 2002, pp. 311–318
2002
Earlier work this paper cites.
P. Newbold, W. L. Carlson, and B. M. Thorne, Statistics for business and economics . Pearson, 2013
2013
Earlier work this paper cites.
J. Hardin, R. Hoerl, N. J. Horton, D. Nolan, B. Baumer, O. Hall-Holt, P. Murrell, R. Peng, P. Roback, D. Temple Lang et al. , “Data science in statistics curricula: Preparing students to “think with data”,” The American Statistician , vol. 69, no. 4, pp. 343–353, 2015
2015
Earlier work this paper cites.
D. Donoho, “50 years of data science,” Journal of Computational and Graphical Statistics , vol. 26, no. 4, pp. 745–766, 2017
2017
Earlier work this paper cites.
Y. Luo, X. Qin, N. Tang, and G. Li, “Deepeye: Towards automatic data visualization,” in ICDE . IEEE Computer Society, 2018, pp. 101–112
2018
Earlier work this paper cites.
K. Li and G. Li, “Approximate query processing: What is new and where to go? a survey on approximate query processing,” Data Science and Engineering , vol. 3, no. 4, pp. 379–397, 2018
2018
Earlier work this paper cites.
F. Emmert-Streib and M. Dehmer, “Understanding statistical hypothesis testing: The logic of statistical inference,” Machine Learning and Knowledge Extraction , vol. 1, no. 3, pp. 945–962, 2019
2019
Earlier work this paper cites.
J. Fan, R. Li, C.-H. Zhang, and H. Zou, Statistical foundations of data science . Chapman and Hall/CRC, 2020
2020
Earlier work this paper cites.
X. Qin, Y. Luo, N. Tang, and G. Li, “Making data visualization more efficient and effective: a survey,” VLDB J. , vol. 29, no. 1, pp. 93–117, 2020
2020
Earlier work this paper cites.
Y. Luo, C. Chai, X. Qin, N. Tang, and G. Li, “Visclean: Interactive cleaning for progressive visualization,” Proc. VLDB Endow. , vol. 13, no. 12, pp. 2821–2824, 2020
2020
Earlier work this paper cites.
T. Zhang, V. Kishore, F. Wu, K. Q. Weinberger, and Y. Artzi, “Bertscore: Evaluating text generation with bert,” in International Conference on Learning Representations , 2020. [Online]. Available: https://openreview.net/forum?id=SkeHuCVFDr
2020
Earlier work this paper cites.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell et al. , “Language models are few-shot learners,” Advances in neural information processing systems , vol. 33, pp. 1877–1901, 2020
2020
Earlier work this paper cites.
Y. Luo, N. Tang, G. Li, W. Li, T. Zhao, and X. Yu, “Deepeye: A data science system for monitoring and exploring COVID-19 data,” IEEE Data Eng. Bull. , vol. 43, no. 2, pp. 121–132, 2020
2020
Earlier work this paper cites.
Y. Luo, C. Chai, X. Qin, N. Tang, and G. Li, “Interactive cleaning for progressive visualization through composite questions,” in ICDE . IEEE, 2020, pp. 733–744
2020
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
D. Hendrycks, C. Burns, S. Basart, A. Zou, M. Mazeika, D. Song, and J. Steinhardt, “Measuring massive multitask language understanding,” Proceedings of the International Conference on Learning Representations (ICLR) , 2021
2021
Earlier work this paper cites.
Y. Luo, N. Tang, G. Li, C. Chai, W. Li, and X. Qin, “Synthesizing natural language to visualization (NL2VIS) benchmarks from NL2SQL benchmarks,” in SIGMOD Conference . ACM, 2021, pp. 1235–1247
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
T. Gebru, J. Morgenstern, B. Vecchione, J. W. Vaughan, H. Wallach, H. D. Iii, and K. Crawford, “Datasheets for datasets,” Communications of the ACM , vol. 64, no. 12, pp. 86–92, 2021
2021
Earlier work this paper cites.
L. Ouyang, J. Wu, X. Jiang, D. Almeida, C. Wainwright, P. Mishkin, C. Zhang, S. Agarwal, K. Slama, A. Ray et al. , “Training language models to follow instructions with human feedback,” Advances in neural information processing systems , vol. 35, pp. 27 730–27 744, 2022
2022
Cited alongside, same era.
J. Wei, X. Wang, D. Schuurmans, M. Bosma, F. Xia, E. Chi, Q. V. Le, D. Zhou et al. , “Chain-of-thought prompting elicits reasoning in large language models,” Advances in neural information processing systems , vol. 35, pp. 24 824–24 837, 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
L. Shen, E. Shen, Y. Luo, X. Yang, X. Hu, X. Zhang, Z. Tai, and J. Wang, “Towards natural language interfaces for data visualization: A survey,” IEEE Trans. Vis. Comput. Graph. , vol. 29, no. 6, pp. 3121–3144, 2023
2023
Later among the works it cites.
C. Chai, J. Wang, Y. Luo, Z. Niu, and G. Li, “Data management for machine learning: A survey,” IEEE Trans. Knowl. Data Eng. , vol. 35, no. 5, pp. 4646–4667, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2022
Cited alongside, same era.
S. Mishra, M. Finlayson, P. Lu, L. Tang, S. Welleck, C. Baral, T. Rajpurohit, O. Tafjord, A. Sabharwal, P. Clark, and A. Kalyan, “Lila: A unified benchmark for mathematical reasoning,” in Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing (EMNLP) , 2022
2022
Cited alongside, same era.
Y. Luo, N. Tang, G. Li, J. Tang, C. Chai, and X. Qin, “Natural language to visualization by neural machine translation,” IEEE Trans. Vis. Comput. Graph. , vol. 28, no. 1, pp. 217–226, 2022
2022
Cited alongside, same era.
J. Tang, Y. Luo, M. Ouzzani, G. Li, and H. Chen, “Sevi: Speech-to-visualization through neural machine translation,” in SIGMOD Conference . ACM, 2022, pp. 2353–2356
2022
Cited alongside, same era.
Y. Luo, X. Qin, C. Chai, N. Tang, G. Li, and W. Li, “Steerable self-driving data visualization,” IEEE Trans. Knowl. Data Eng. , vol. 34, no. 1, pp. 475–490, 2022
2022
Cited alongside, same era.
H. Hassani and E. S. Silva, “The role of chatgpt in data science: how ai-assisted conversational interfaces are revolutionizing the field,” Big data and cognitive computing , vol. 7, no. 2, p. 62, 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Later among the works it cites.
W. Kwon, Z. Li, S. Zhuang, Y. Sheng, L. Zheng, C. H. Yu, J. E. Gonzalez, H. Zhang, and I. Stoica, “Efficient memory management for large language model serving with pagedattention,” in Proceedings of the ACM SIGOPS 29th Symposium on Operating Systems Principles , 2023
2023
Later among the works it cites.
2024
Closest in time.
G. Li, R. Li, Y. Feng, Y. Zhang, Y. Luo, and C. H. Liu, “Coinsight: Visual storytelling for hierarchical tables with connected insights,” IEEE Trans. Vis. Comput. Graph. , vol. 30, no. 6, pp. 3049–3061, 2024
2024
Closest in time.
Y. Zhuang, Y. Yu, K. Wang, H. Sun, and C. Zhang, “Toolqa: A dataset for llm question answering with external tools,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
Y. Ye, J. Hao, Y. Hou, Z. Wang, S. Xiao, Y. Luo, and W. Zeng, “Generative AI for visualization: State of the art and future directions,” Vis. Informatics , vol. 8, no. 1, pp. 43–66, 2024
2024
Closest in time.
N. Tang, C. Yang, J. Fan, L. Cao, Y. Luo, and A. Y. Halevy, “Verifai: Verified generative AI,” in CIDR . www.cidrdb.org, 2024
2024
Closest in time.
P. Lu, H. Bansal, T. Xia, J. Liu, C. Li, H. Hajishirzi, H. Cheng, K.-W. Chang, M. Galley, and J. Gao, “Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts,” in International Conference on Learning Representations (ICLR) , 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
X. Hu, Z. Zhao, S. Wei, Z. Chai, Q. Ma, G. Wang, X. Wang, J. Su, J. Xu, M. Zhu, Y. Cheng, J. Yuan, J. Li, K. Kuang, Y. Yang, H. Yang, and F. Wu, “Infiagent-dabench: Evaluating agents on data analysis tasks,” 2024
2024
Closest in time.
V. Arel-Bundock, Rdatasets: A collection of datasets originally distributed in various R packages , 2024, r package version 1.0.0. [Online]. Available: https://vincentarelbundock.github.io/Rdatasets
2024
Closest in time.
B. Li, Y. Luo, C. Chai, G. Li, and N. Tang, “The dawn of natural language to SQL: are we fully ready?” Proc. VLDB Endow. , vol. 17, no. 11, pp. 3318–3331, 2024
2024
Closest in time.
S. Frieder, L. Pinchetti, R.-R. Griffiths, T. Salvatori, T. Lukasiewicz, P. Petersen, and J. Berner, “Mathematical capabilities of chatgpt,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
Y. Xie, Y. Luo, G. Li, and N. Tang, “Haichart: Human and AI paired visualization system,” Proc. VLDB Endow. , vol. 17, no. 11, pp. 3178–3191, 2024
2024
Closest in time.
Y. Chang, X. Wang, J. Wang, Y. Wu, L. Yang, K. Zhu, H. Chen, X. Yi, C. Wang, Y. Wang et al. , “A survey on evaluation of large language models,” ACM Transactions on Intelligent Systems and Technology , vol. 15, no. 3, pp. 1–45, 2024
2024
Closest in time.