Fetching the paper…
Reading the bibliography…
Artificial Intelligence (AI) has increasingly influenced modern society, recently in particular through significant advancements in Large Language Models (LLMs).
1903
Earlier work this paper cites.
1905
Earlier work this paper cites.
1909
Earlier work this paper cites.
A. Vargha, H. D. Delaney, and A. Vargha, “A Critique and Improvement of the ”CL” Common Language Effect Size Statistics of McGraw and Wong,” Journal of Educational and Behavioral Statistics , vol. 25, no. 2, p. 101, 2000. [Online]. Available: http://links.jstor.org/sici?sici=1076-9986%28200022%2925%3A2%3C101%3AACAIOT%3E2.0.CO%3B2-O&origin=crossref
2000
Earlier work this paper cites.
2003
Earlier work this paper cites.
H. O. Mayer, Interview und schriftliche Befragung: Entwicklung, Durchführung und Auswertung , 4th ed., ser. 150 Jahre Wissen für die Zukunft. München Wien: Oldenbourg, 2008
2008
Earlier work this paper cites.
G. Norman, “Likert scales, levels of measurement and the “laws” of statistics,” Advances in Health Sciences Education , vol. 15, no. 5, pp. 625–632, Dec. 2010. [Online]. Available: http://link.springer.com/10.1007/s10459-010-9222-y
2010
Earlier work this paper cites.
A. Bhattacherjee, Social Science Research: Principles, Methods and Practices . Open Textbook Library, 01 2012
2012
Earlier work this paper cites.
2015
Earlier work this paper cites.
A. Dinno, “Nonparametric Pairwise Multiple Comparisons in Independent Groups using Dunn’s Test,” The Stata Journal: Promoting communications on statistics and Stata , vol. 15, no. 1, pp. 292–300, Apr. 2015. [Online]. Available: http://journals.sagepub.com/doi/10.1177/1536867X1501500117
2015
Earlier work this paper cites.
L. A. Hendricks, Z. Akata, M. Rohrbach, J. Donahue, B. Schiele, and T. Darrell, “Generating visual explanations,” in Computer Vision – ECCV 2016 , B. Leibe, J. Matas, N. Sebe, and M. Welling, Eds. Cham: Springer International Publishing, 2016, pp. 3–19
2016
Earlier work this paper cites.
F. Doshi-Velez and B. Kim, “Towards a rigorous science of interpretable machine learning,” arXiv: Machine Learning , 2017. [Online]. Available: https://api.semanticscholar.org/CorpusID:11319376
2017
Earlier work this paper cites.
G. Chen, W. Choi, X. Yu, T. Han, and M. Chandraker, “Learning Efficient Object Detection Models with Knowledge Distillation,” in Advances in Neural Information Processing Systems , vol. 30. Curran Associates, Inc., 2017. [Online]. Available: https://papers.nips.cc/paper_files/paper/2017/hash/e1e32e235eee1f970470a3a6658dfdd5-Abstract.html
2017
Earlier work this paper cites.
J. Kim, A. Rohrbach, T. Darrell, J. Canny, and Z. Akata, “Textual explanations for self-driving vehicles,” in Computer Vision – ECCV 2018 , V. Ferrari, M. Hebert, C. Sminchisescu, and Y. Weiss, Eds. Cham: Springer International Publishing, 2018, pp. 577–593
2018
Earlier work this paper cites.
D. H. Park, L. A. Hendricks, Z. Akata, A. Rohrbach, B. Schiele, T. Darrell, and M. Rohrbach, “Multimodal explanations: Justifying decisions and pointing to the evidence,” 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition , pp. 8779–8788, 2018. [Online]. Available: https://api.semanticscholar.org/CorpusID:3604848
2018
Earlier work this paper cites.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, I. Sutskever et al. , “Language models are unsupervised multitask learners,” OpenAI blog , vol. 1, no. 8, p. 9, 2019
2019
Earlier work this paper cites.
P. Izsak, S. Guskin, and M. Wasserblat, “Training Compact Models for Low Resource Entity Tagging using Pre-trained Language Models,” in 2019 Fifth Workshop on Energy Efficient Machine Learning and Cognitive Computing - NeurIPS Edition (EMC2-NIPS) , Dec. 2019, pp. 44–47. [Online]. Available: https://ieeexplore.ieee.org/document/9463575#page=1.73
2019
Earlier work this paper cites.
A. Holzinger, G. Langs, H. Denk, K. Zatloukal, and H. Müller, “Causability and explainability of artificial intelligence in medicine,” Wiley Interdisciplinary Reviews. Data Mining and Knowledge Discovery , vol. 9, 2019. [Online]. Available: https://api.semanticscholar.org/CorpusID:132372213
2019
Earlier work this paper cites.
N. F. Rajani, B. McCann, C. Xiong, and R. Socher, “Explain Yourself! Leveraging Language Models for Commonsense Reasoning,” in Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics . Florence, Italy: Association for Computational Linguistics, 2019, pp. 4932–4942. [Online]. Available: https://www.aclweb.org/anthology/P19-1487
2019
Earlier work this paper cites.
A. Talmor, J. Herzig, N. Lourie, and J. Berant, “CommonsenseQA: A Question Answering Challenge Targeting Commonsense Knowledge,” in Proceedings of the 2019 Conference of the North . Minneapolis, Minnesota: Association for Computational Linguistics, 2019, pp. 4149–4158. [Online]. Available: http://aclweb.org/anthology/N19-1421
2019
Earlier work this paper cites.
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, S. Agarwal, A. Herbert-Voss, G. Krueger, T. Henighan, R. Child, A. Ramesh, D. M. Ziegler, J. Wu, C. Winter, C. Hesse, M. Chen, E. Sigler, M. Litwin, S. Gray, B. Chess, J. Clark, C. Berner, S. McCandlish, A. Radford, I. Sutskever, and D. Amodei, “Language models are few-shot learners,” in Proceedings of the 34th International Conference on Neural Information Processing Systems , ser. NIPS ’20. Red Hook, NY, USA: Curran Associates Inc., 2020
2020
Earlier work this paper cites.
X. Jiao, Y. Yin, L. Shang, X. Jiang, X. Chen, L. Li, F. Wang, and Q. Liu, “TinyBERT: Distilling BERT for Natural Language Understanding,” in Findings of the Association for Computational Linguistics: EMNLP 2020 , T. Cohn, Y. He, and Y. Liu, Eds. Online: Association for Computational Linguistics, Nov. 2020, pp. 4163–4174. [Online]. Available: https://aclanthology.org/2020.findings-emnlp.372/
2020
Earlier work this paper cites.
E. Strubell, A. Ganesh, and A. McCallum, “Energy and policy considerations for modern deep learning research,” Proceedings of the AAAI Conference on Artificial Intelligence , vol. 34, no. 09, pp. 13 693–13 696, Apr. 2020. [Online]. Available: https://ojs.aaai.org/index.php/AAAI/article/view/7123
2020
Earlier work this paper cites.
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu, “Exploring the limits of transfer learning with a unified text-to-text transformer,” J. Mach. Learn. Res. , vol. 21, no. 1, Jan. 2020
2020
Earlier work this paper cites.
T. Wolf, L. Debut, V. Sanh, J. Chaumond, C. Delangue, A. Moi, P. Cistac, T. Rault, R. Louf, M. Funtowicz, J. Davison, S. Shleifer, P. von Platen, C. Ma, Y. Jernite, J. Plu, C. Xu, T. Le Scao, S. Gugger, M. Drame, Q. Lhoest, and A. Rush, “Transformers: State-of-the-art natural language processing,” in Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: System Demonstrations , Q. Liu and D. Schlangen, Eds. Online: Association for Computational Linguistics, Oct. 2020, pp. 38–45. [Online]. Available: https://aclanthology.org/2020.emnlp-demos.6/
2020
Earlier work this paper cites.
P. Hase, S. Zhang, H. Xie, and M. Bansal, “Leakage-adjusted simulatability: Can models generate non-trivial explanations of their behavior in natural language?” in Findings of the Association for Computational Linguistics: EMNLP 2020 , T. Cohn, Y. He, and Y. Liu, Eds. Online: Association for Computational Linguistics, Nov. 2020, pp. 4351–4367. [Online]. Available: https://aclanthology.org/2020.findings-emnlp.390/
2020
Earlier work this paper cites.
A. Holzinger, A. Carrington, and H. Müller, “Measuring the Quality of Explanations: The System Causability Scale (SCS): Comparing Human and Machine Explanations,” KI - Künstliche Intelligenz , vol. 34, no. 2, pp. 193–198, Jun. 2020. [Online]. Available: http://link.springer.com/10.1007/s13218-020-00636-z
2020
Earlier work this paper cites.
T. Le Scao and A. Rush, “How many data points is a prompt worth?” in Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . Online: Association for Computational Linguistics, 2021, pp. 2627–2636. [Online]. Available: https://aclanthology.org/2021.naacl-main.208
2021
Earlier work this paper cites.
2021
Cited alongside, same era.
C. Wu, F. Wu, and Y. Huang, “One teacher is enough? pre-trained language model distillation from multiple teachers,” in Findings of the Association for Computational Linguistics: ACL-IJCNLP 2021 , C. Zong, F. Xia, W. Li, and R. Navigli, Eds. Online: Association for Computational Linguistics, Aug. 2021, pp. 4408–4413. [Online]. Available: https://aclanthology.org/2021.findings-acl.387/
2021
Cited alongside, same era.
R. He, L. Liu, H. Ye, Q. Tan, B. Ding, L. Cheng, J. Low, L. Bing, and L. Si, “On the effectiveness of adapter-based tuning for pretrained language model adaptation,” in Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) . Association for Computational Linguistics, 2021. [Online]. Available: http://dx.doi.org/10.18653/v1/2021.acl-long.172
X. Wang, J. Wei, D. Schuurmans, Q. V. Le, E. H. Chi, S. Narang, A. Chowdhery, and D. Zhou, “Self-consistency improves chain of thought reasoning in language models,” in The Eleventh International Conference on Learning Representations , 2023. [Online]. Available: https://openreview.net/forum?id=1PL1NIMMrw
2023
Later among the works it cites.
2023
Later among the works it cites.
Prolific, “Prolific · Quickly find research participants you can trust.” 2023. [Online]. Available: https://www.prolific.com/
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2021
Cited alongside, same era.
2021
Cited alongside, same era.
B. Lester, R. Al-Rfou, and N. Constant, “The Power of Scale for Parameter-Efficient Prompt Tuning,” in Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing , M.-F. Moens, X. Huang, L. Specia, and S. W.-t. Yih, Eds. Online and Punta Cana, Dominican Republic: Association for Computational Linguistics, Nov. 2021, pp. 3045–3059. [Online]. Available: https://aclanthology.org/2021.emnlp-main.243/
2021
Cited alongside, same era.
N. Burkart and M. F. Huber, “A survey on the explainability of supervised machine learning,” Journal of Artificial Intelligence Research , vol. 70, pp. 245–317, 2021
2021
Cited alongside, same era.
S. Wiegreffe, A. Marasović, and N. A. Smith, “Measuring Association Between Labels and Free-Text Rationales,” in Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing , M.-F. Moens, X. Huang, L. Specia, and S. W.-t. Yih, Eds. Online and Punta Cana, Dominican Republic: Association for Computational Linguistics, Nov. 2021, pp. 10 266–10 284. [Online]. Available: https://aclanthology.org/2021.emnlp-main.804/
2021
Cited alongside, same era.
J. Wei, X. Wang, D. Schuurmans, M. Bosma, B. Ichter, F. Xia, E. H. Chi, Q. V. Le, and D. Zhou, “Chain-of-thought prompting elicits reasoning in large language models,” in Proceedings of the 36th International Conference on Neural Information Processing Systems , ser. NIPS ’22. Red Hook, NY, USA: Curran Associates Inc., 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
M. Heikkilä, “We’re getting a better idea of AI’s true carbon footprint,” Nov. 2022. [Online]. Available: https://www.technologyreview.com/2022/11/14/1063192/were-getting-a-better-idea-of-ais-true-carbon-footprint/
2022
Cited alongside, same era.
2022
Cited alongside, same era.
S. Black, S. Biderman, E. Hallahan, Q. Anthony, L. Gao, L. Golding, H. He, C. Leahy, K. McDonell, J. Phang, M. Pieler, U. S. Prashanth, S. Purohit, L. Reynolds, J. Tow, B. Wang, and S. Weinbach, “GPT-NeoX-20B: An open-source autoregressive language model,” in Proceedings of BigScience Episode #5 – Workshop on Challenges & Perspectives in Creating Large Language Models , A. Fan, S. Ilic, T. Wolf, and M. Gallé, Eds. virtual+Dublin: Association for Computational Linguistics, May 2022, pp. 95–136. [Online]. Available: https://aclanthology.org/2022.bigscience-1.9/
2022
Cited alongside, same era.
H. Chen, F. Brahman, X. Ren, Y. Ji, Y. Choi, and S. Swayamdipta, “REV: Information-theoretic evaluation of free-text rationales,” in Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , A. Rogers, J. Boyd-Graber, and N. Okazaki, Eds. Toronto, Canada: Association for Computational Linguistics, Jul. 2023, pp. 2007–2030. [Online]. Available: https://aclanthology.org/2023.acl-long.112/
2023
Later among the works it cites.
R. R. Hoffman, S. T. Mueller, G. Klein, and J. Litman, “Measures for explainable AI: Explanation goodness, user satisfaction, mental models, curiosity, trust, and human-AI performance,” Frontiers in Computer Science , vol. 5, p. 1096257, Feb. 2023. [Online]. Available: https://www.frontiersin.org/articles/10.3389/fcomp.2023.1096257/full
2023
Later among the works it cites.
2023
Later among the works it cites.
2024
Later among the works it cites.
M. Parmar, N. Patel, N. Varshney, M. Nakamura, M. Luo, S. Mashetty, A. Mitra, and C. Baral, “LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models,” in Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , L.-W. Ku, A. Martins, and V. Srikumar, Eds. Bangkok, Thailand: Association for Computational Linguistics, Aug. 2024, pp. 13 679–13 707. [Online]. Available: https://aclanthology.org/2024.acl-long.739/
2024
Later among the works it cites.
2024
Later among the works it cites.
X. Zhu, J. Li, Y. Liu, C. Ma, and W. Wang, “A survey on model compression for large language models,” Transactions of the Association for Computational Linguistics , vol. 12, pp. 1556–1577, 2024. [Online]. Available: https://aclanthology.org/2024.tacl-1.85/
2024
Later among the works it cites.
S. Kim, G. Ham, Y. Cho, and D. Kim, “Robustness-Reinforced Knowledge Distillation With Correlation Distance and Network Pruning,” IEEE Transactions on Knowledge and Data Engineering , vol. 36, no. 12, pp. 9163–9175, Dec. 2024, conference Name: IEEE Transactions on Knowledge and Data Engineering. [Online]. Available: https://ieeexplore.ieee.org/document/10623281/?arnumber=10623281
2024
Later among the works it cites.
Q. Zhong, L. Ding, J. Liu, B. Du, and D. Tao, “PanDa: Prompt Transfer Meets Knowledge Distillation for Efficient Model Adaptation,” IEEE Transactions on Knowledge and Data Engineering , vol. 36, no. 9, pp. 4835–4848, Sep. 2024, conference Name: IEEE Transactions on Knowledge and Data Engineering. [Online]. Available: https://ieeexplore.ieee.org/document/10475529/?arnumber=10475529
2024
Later among the works it cites.
2024
Later among the works it cites.
C. Dai, K. Li, W. Zhou, and S. Hu, “Improve Student‘s Reasoning Generalizability through Cascading Decomposed CoTs Distillation,” in Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , Y. Al-Onaizan, M. Bansal, and Y.-N. Chen, Eds. Miami, Florida, USA: Association for Computational Linguistics, Nov. 2024, pp. 15 623–15 643. [Online]. Available: https://aclanthology.org/2024.emnlp-main.875/
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Y. Gu, L. Dong, F. Wei, and M. Huang, “MiniLLM: Knowledge distillation of large language models,” in The Twelfth International Conference on Learning Representations , 2024. [Online]. Available: https://openreview.net/forum?id=5h0qf7IBZZ
2024
Later among the works it cites.
H. Yu, R. Li, Z. Zhang, S. Ye, Q. Liu, Z. Huang, and E. Chen, “ERDL: Efficient Retrieval Framework Based on Distillation from Large Language Models,” in 2024 International Conference on Computational Linguistics and Natural Language Processing (CLNLP) , Jul. 2024, pp. 83–87. [Online]. Available: https://ieeexplore.ieee.org/document/10743156
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
T. Mei, Y. Zi, X. Cheng, Z. Gao, Q. Wang, and H. Yang, “Efficiency Optimization of Large-Scale Language Models Based on Deep Learning in Natural Language Processing Tasks,” in 2024 IEEE 2nd International Conference on Sensors, Electronics and Computer Engineering (ICSECE) , Aug. 2024, pp. 1231–1237. [Online]. Available: https://ieeexplore.ieee.org/document/10729518/?arnumber=10729518
2024
Later among the works it cites.
K. Yang, D. Klein, A. Celikyilmaz, N. Peng, and Y. Tian, “RLCD: Reinforcement learning from contrastive distillation for LM alignment,” in The Twelfth International Conference on Learning Representations , 2024. [Online]. Available: https://openreview.net/forum?id=v3XXtxWKi6
2024
Later among the works it cites.
O. Dictionary, “Oxford Learner’s Dictionaries | Find definitions, translations, and grammar explanations at Oxford Learner’s Dictionaries,” 2024. [Online]. Available: https://www.oxfordlearnersdictionaries.com/
2024
Later among the works it cites.
N. Alangari, M. El Bachir Menai, H. Mathkour, and I. Almosallam, “Exploring Evaluation Methods for Interpretable Machine Learning: A Survey,” Information , vol. 14, no. 8, p. 469, Aug. 2023. [Online]. Available: https://www.mdpi.com/2078-2489/14/8/469
2078
Closest in time.
J. Zhou, A. H. Gandomi, F. Chen, and A. Holzinger, “Evaluating the Quality of Machine Learning Explanations: A Survey on Methods and Metrics,” Electronics , vol. 10, no. 5, p. 593, Mar. 2021. [Online]. Available: https://www.mdpi.com/2079-9292/10/5/593
2079
Closest in time.
D. V. Carvalho, E. M. Pereira, and J. S. Cardoso, “Machine Learning Interpretability: A Survey on Methods and Metrics,” Electronics , vol. 8, no. 8, p. 832, Jul. 2019. [Online]. Available: https://www.mdpi.com/2079-9292/8/8/832
2079
Closest in time.