L. Perlovsky, “Language and emotions: emotional sapir–whorf hypothesis,” Neural Networks , vol. 22, no. 5-6, pp. 518–526, 2009
2009
Earlier work this paper cites.
R. Liu and N. Moini, “Benchmarking transportation safety performance via shift-share approaches,” Journal of Transportation Safety & Security , vol. 7, no. 2, pp. 124–137, 2015
2015
Earlier work this paper cites.
M. I. Javaid and M. M. W. Iqbal, “A comprehensive people, process and technology (ppt) application model for information systems (is) risk management in small/medium enterprises (sme),” in 2017 International Conference on Communication Technologies (ComTech) . IEEE, 2017, pp. 78–90
2017
Earlier work this paper cites.
B. Li, W. Ren, D. Fu, D. Tao, D. Feng, W. Zeng, and Z. Wang, “Benchmarking single-image dehazing and beyond,” IEEE Transactions on Image Processing , vol. 28, no. 1, pp. 492–505, 2018
2018
Earlier work this paper cites.
A. V. Papadopoulos, L. Versluis, A. Bauer, N. Herbst, J. Von Kistowski, A. Ali-Eldin, C. L. Abad, J. N. Amaral, P. Tuma, and A. Iosup, “Methodological principles for reproducible performance evaluation in cloud computing,” IEEE Transactions on Software Engineering , vol. 47, no. 8, pp. 1528–1543, 2019
2019
Earlier work this paper cites.
A. Wang, A. Singh, J. Michael, F. Hill, O. Levy, and S. R. Bowman, “Glue: A multi-task benchmark and analysis platform for natural language understanding,” in 7th International Conference on Learning Representations, ICLR 2019 , 2019
2019
Earlier work this paper cites.
M. S. Aslanpour, S. S. Gill, and A. N. Toosi, “Performance evaluation metrics for cloud, fog and edge computing: A review, taxonomy, benchmarks and standards for future research,” Internet of Things , vol. 12, p. 100273, 2020
2020
Earlier work this paper cites.
F. Xia, W. B. Shen, C. Li, P. Kasimbeg, M. E. Tchapmi, A. Toshev, R. Martín-Martín, and S. Savarese, “Interactive gibson benchmark: A benchmark for interactive navigation in cluttered environments,” IEEE Robotics and Automation Letters , vol. 5, no. 2, pp. 713–720, 2020
2020
Earlier work this paper cites.
S. Zhu, J. Shi, L. Yang, B. Qin, Z. Zhang, L. Song, and G. Wang, “Measuring and modeling the label dynamics of online { \{ Anti-Malware } \} engines,” in 29th USENIX Security Symposium (USENIX Security 20) , 2020, pp. 2361–2378
2020
Earlier work this paper cites.
D. Al-Fraihat, M. Joy, J. Sinclair et al. , “Evaluating e-learning systems success: An empirical study,” Computers in human behavior , vol. 102, pp. 67–86, 2020
2020
Earlier work this paper cites.
D. Hendrycks, C. Burns, S. Basart, A. Zou, M. Mazeika, D. Song, and J. Steinhardt, “Measuring massive multitask language understanding,” arXiv preprint arXiv:2009.03300 , 2020
Original
2020
Earlier work this paper cites.
M. Chen, J. Tworek, H. Jun, Q. Yuan, H. P. d. O. Pinto, J. Kaplan, H. Edwards, Y. Burda, N. Joseph, G. Brockman et al. , “Evaluating large language models trained on code,” arXiv preprint arXiv:2107.03374 , 2021
Original
2021
Earlier work this paper cites.
M. Hort, M. Kechagia, F. Sarro, and M. Harman, “A survey of performance optimization for mobile applications,” IEEE Transactions on Software Engineering , vol. 48, no. 8, pp. 2879–2904, 2021
2021
Earlier work this paper cites.
A. Romano, X. Liu, Y. Kwon, and W. Wang, “An empirical study of bugs in webassembly compilers,” in 2021 36th IEEE/ACM International Conference on Automated Software Engineering (ASE) . IEEE, 2021, pp. 42–54
2021
Earlier work this paper cites.
T. McIntosh, A. Kayes, Y.-P. P. Chen, A. Ng, and P. Watters, “Ransomware mitigation in the modern era: A comprehensive review, research challenges, and future directions,” ACM Computing Surveys (CSUR) , vol. 54, no. 9, pp. 1–36, 2021
2021
Earlier work this paper cites.
A. Srivastava, A. Rastogi, A. Rao, A. A. M. Shoeb, A. Abid, A. Fisch, A. R. Brown, A. Santoro, A. Gupta, A. Garriga-Alonso et al. , “Beyond the imitation game: Quantifying and extrapolating the capabilities of language models,” Transactions on machine learning research , 2022
2022
Earlier work this paper cites.
H. Brunst, S. Chandrasekaran, F. M. Ciorba, N. Hagerty, R. Henschel, G. Juckeland, J. Li, V. G. M. Vergara, S. Wienke, and M. Zavala, “First experiences in performance benchmarking with the new spechpc 2021 suites,” in 2022 22nd IEEE International Symposium on Cluster, Cloud and Internet Computing (CCGrid) . IEEE, 2022, pp. 675–684
2022
Earlier work this paper cites.
P. Liang, R. Bommasani, T. Lee, D. Tsipras, D. Soylu, M. Yasunaga, Y. Zhang, D. Narayanan, Y. Wu, A. Kumar et al. , “Holistic evaluation of language models,” arXiv preprint arXiv:2211.09110 , 2022
Original
2022
Earlier work this paper cites.
M. Jegorova, C. Kaul, C. Mayor, A. Q. O’Neil, A. Weir, R. Murray-Smith, and S. A. Tsaftaris, “Survey: Leakage and privacy at inference time,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 45, no. 7, pp. 9090–9108, 2022
2022
Earlier work this paper cites.
R. Shah, K. Chawla, D. Eidnani, A. Shah, W. Du, S. Chava, N. Raman, C. Smiley, J. Chen, and D. Yang, “When flue meets flang: Benchmarks and large pretrained language model for financial domain,” in Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing , 2022, pp. 2322–2335
2022
Earlier work this paper cites.
B. Min, H. Ross, E. Sulem, A. P. B. Veyseh, T. H. Nguyen, O. Sainz, E. Agirre, I. Heintz, and D. Roth, “Recent advances in natural language processing via large pre-trained language models: A survey,” ACM Computing Surveys , vol. 56, no. 2, pp. 1–40, 2023
2023
Earlier work this paper cites.