Fetching the paper…
Reading the bibliography…
The recent performance leap of Large Language Models (LLMs) opens up new opportunities across numerous industrial applications and domains.
1908
Earlier work this paper cites.
T. Kuhn, H. Niemann, and E. G. Schukat-Talamazzini, “Ergodic hidden markov models and polygrams for language modeling,” in Proceedings of ICASSP’94. IEEE International Conference on Acoustics, Speech and Signal Processing , vol. 1. IEEE, 1994, pp. I–357
1994
Earlier work this paper cites.
D. Barber and C. M. Bishop, “Ensemble learning in bayesian neural networks,” Nato ASI Series F Computer and Systems Sciences , vol. 168, pp. 215–238, 1998
1998
Earlier work this paper cites.
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu, “Bleu: a method for automatic evaluation of machine translation,” in Proceedings of the 40th annual meeting of the Association for Computational Linguistics , 2002, pp. 311–318
2002
Earlier work this paper cites.
C.-Y. Lin, “ROUGE: A package for automatic evaluation of summaries,” in Text Summarization Branches Out . Barcelona, Spain: Association for Computational Linguistics, Jul. 2004, pp. 74–81. [Online]. Available: https://aclanthology.org/W04-1013
2004
Earlier work this paper cites.
T. Brants, A. Popat, P. Xu, F. J. Och, and J. Dean, “Large language models in machine translation,” in Proceedings of the 2007 Joint Conference on Empirical Methods in Natural Language Processing and Computational Natural Language Learning (EMNLP-CoNLL) , 2007, pp. 858–867
2007
Earlier work this paper cites.
T. Mikolov, M. Karafiát, L. Burget, J. Cernockỳ, and S. Khudanpur, “Recurrent neural network based language model.” in Interspeech , vol. 2, no. 3. Makuhari, 2010, pp. 1045–1048
2010
Earlier work this paper cites.
M. Barreno, B. Nelson, A. D. Joseph, and J. D. Tygar, “The security of machine learning,” Machine Learning , vol. 81, pp. 121–148, 2010
2010
Earlier work this paper cites.
O. Bojar, C. Buck, C. Federmann, B. Haddow, P. Koehn, J. Leveling, C. Monz, P. Pecina, M. Post, H. Saint-Amand, R. Soricut, L. Specia, and A. Tamchyna, “Findings of the 2014 workshop on statistical machine translation,” in Proceedings of the Ninth Workshop on Statistical Machine Translation . Baltimore, Maryland, USA: Association for Computational Linguistics, Jun. 2014, pp. 12–58. [Online]. Available: https://aclanthology.org/W14-3302
2014
Earlier work this paper cites.
Y. Gal and Z. Ghahramani, “Dropout as a bayesian approximation: Representing model uncertainty in deep learning,” in international conference on machine learning . PMLR, 2016, pp. 1050–1059
2016
Earlier work this paper cites.
Y. Gal, “Uncertainty in deep learning.” University of Cambridge., 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
K. R. Varshney and H. Alemzadeh, “On the safety of machine learning: Cyber-physical systems, decision sciences, and data products,” Big data , vol. 5, no. 3, pp. 246–255, 2017
2017
Earlier work this paper cites.
A. Kendall and Y. Gal, “What uncertainties do we need in bayesian deep learning for computer vision?” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
B. Lakshminarayanan, A. Pritzel, and C. Blundell, “Simple and scalable predictive uncertainty estimation using deep ensembles,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
D. Hendrycks and K. Gimpel, “A baseline for detecting misclassified and out-of-distribution examples in neural networks,” in 5th International Conference on Learning Representations, ICLR 2017, Toulon, France, April 24-26, 2017, Conference Track Proceedings . OpenReview.net, 2017. [Online]. Available: https://openreview.net/forum?id=Hkg4TI9xl
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
P. Oberdiek, M. Rottmann, and H. Gottschalk, “Classification uncertainty of deep neural networks based on gradient information,” in Artificial Neural Networks in Pattern Recognition: 8th IAPR TC3 Workshop, ANNPR 2018, Siena, Italy, September 19–21, 2018, Proceedings 8 . Springer, 2018, pp. 113–125
2018
Earlier work this paper cites.
Y. Yang, W.-t. Yih, and C. Meek, “WikiQA: A challenge dataset for open-domain question answering,” in Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing . Lisbon, Portugal: Association for Computational Linguistics, Sep. 2015, pp. 2013–2018. [Online]. Available: https://aclanthology.org/D15-1237
2018
Earlier work this paper cites.
O. Rahmati, B. Choubin, A. Fathabadi, F. Coulon, E. Soltani, H. Shahabi, E. Mollaefar, J. Tiefenbacher, S. Cipullo, B. B. Ahmad et al. , “Predicting uncertainty of machine learning models for modelling nitrate pollution of groundwater using quantile regression and uneec methods,” Science of the Total Environment , vol. 688, pp. 855–866, 2019
2019
Earlier work this paper cites.
J. Kim, R. Feldt, and S. Yoo, “Guiding deep learning system testing using surprise adequacy,” in 2019 IEEE/ACM 41st International Conference on Software Engineering (ICSE) . IEEE, 2019, pp. 1039–1049
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
J. D. M.-W. C. Kenton and L. K. Toutanova, “Bert: Pre-training of deep bidirectional transformers for language understanding,” in Proceedings of NAACL-HLT , 2019, pp. 4171–4186
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, I. Sutskever et al. , “Language models are unsupervised multitask learners,” OpenAI blog , vol. 1, no. 8, p. 9, 2019
2019
Earlier work this paper cites.
S. Rabanser, S. Günnemann, and Z. Lipton, “Failing loudly: An empirical study of methods for detecting dataset shift,” Advances in Neural Information Processing Systems , vol. 32, 2019
2019
Earlier work this paper cites.
C. Ning and F. You, “Optimization under uncertainty in the era of big data and deep learning: When machine learning meets mathematical programming,” Computers & Chemical Engineering , vol. 125, pp. 434–448, 2019
2019
Earlier work this paper cites.
G. Wang, W. Li, M. Aertsen, J. Deprest, S. Ourselin, and T. Vercauteren, “Aleatoric uncertainty estimation with test-time augmentation for medical image segmentation with convolutional neural networks,” Neurocomputing , vol. 338, pp. 34–45, 2019
2019
Earlier work this paper cites.
J. Vig and Y. Belinkov, “Analyzing the structure of attention in a transformer language model,” in Proceedings of the 2019 ACL Workshop BlackboxNLP: Analyzing and Interpreting Neural Networks for NLP , 2019, pp. 63–76
2019
Earlier work this paper cites.
2020
Earlier work this paper cites.
X. Zhang, X. Xie, L. Ma, X. Du, Q. Hu, Y. Liu, J. Zhao, and M. Sun, “Towards characterizing adversarial defects of deep learning software from the lens of uncertainty,” in Proceedings of the ACM/IEEE 42nd International Conference on Software Engineering , 2020, pp. 739–751
2020
Earlier work this paper cites.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell et al. , “Language models are few-shot learners,” Advances in neural information processing systems , vol. 33, pp. 1877–1901, 2020
2020
Earlier work this paper cites.
J. Yang, M. Wang, H. Zhou, C. Zhao, W. Zhang, Y. Yu, and L. Li, “Towards making the most of bert in neural machine translation,” in Proceedings of the AAAI conference on artificial intelligence , vol. 34, no. 05, 2020, pp. 9378–9385
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
M. Lewis, Y. Liu, N. Goyal, M. Ghazvininejad, A. Mohamed, O. Levy, V. Stoyanov, and L. Zettlemoyer, “Bart: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension,” in Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics , 2020, pp. 7871–7880
2020
Earlier work this paper cites.
S. Lo Piano, “Ethical principles in machine learning and artificial intelligence: cases from the field and possible ways forward,” Humanities and Social Sciences Communications , vol. 7, no. 1, pp. 1–7, 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
Y.-C. Hsu, Y. Shen, H. Jin, and Z. Kira, “Generalized odin: Detecting out-of-distribution image without learning from out-of-distribution data,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 10 951–10 960
2020
Earlier work this paper cites.
A. Lyzhov, Y. Molchanova, A. Ashukha, D. Molchanov, and D. Vetrov, “Greedy policy search: A simple baseline for learnable test-time augmentation,” in Conference on Uncertainty in Artificial Intelligence . PMLR, 2020, pp. 1308–1317
2020
Earlier work this paper cites.
A. Galassi, M. Lippi, and P. Torroni, “Attention in natural language processing,” IEEE transactions on neural networks and learning systems , vol. 32, no. 10, pp. 4291–4308, 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
M. T. Ribeiro, T. Wu, C. Guestrin, and S. Singh, “Beyond accuracy: Behavioral testing of NLP models with CheckList,” in Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics , D. Jurafsky, J. Chai, N. Schluter, and J. Tetreault, Eds. Online: Association for Computational Linguistics, Jul. 2020, pp. 4902–4912. [Online]. Available: https://aclanthology.org/2020.acl-main.442
2020
Cited alongside, same era.
——, “Responsible ai,” 2023. [Online]. Available: https://perma.cc/9BAY-NM9D
2023
Closest in time.
——, “Building generative ai features responsibly,” 2023. [Online]. Available: https://about.fb.com/news/2023/09/building-generative-ai-features-responsibly/
2023
Closest in time.
2023
Closest in time.
X. Hou, Y. Zhao, Y. Liu, Z. Yang, K. Wang, L. Li, X. Luo, D. Lo, J. Grundy, and H. Wang, “Large language models for software engineering: A systematic literature review,” 2023
2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2021
Cited alongside, same era.
A. Abid, M. Farooqi, and J. Zou, “Persistent anti-muslim bias in large language models,” in Proceedings of the 2021 AAAI/ACM Conference on AI, Ethics, and Society , 2021, pp. 298–306
2021
Cited alongside, same era.
Meta, “Facebook’s five pillars of responsible ai,” 2021. [Online]. Available: https://ai.meta.com/blog/facebooks-five-pillars-of-responsible-ai/
2021
Cited alongside, same era.
D. Kiela, H. Firooz, A. Mohan, V. Goswami, A. Singh, C. A. Fitzpatrick, P. Bull, G. Lipstein, T. Nelli, R. Zhu et al. , “The hateful memes challenge: Competition report,” in NeurIPS 2020 Competition and Demonstration Track . PMLR, 2021, pp. 344–360
2021
Cited alongside, same era.
U. Bhatt, J. Antorán, Y. Zhang, Q. V. Liao, P. Sattigeri, R. Fogliato, G. Melançon, R. Krishnan, J. Stanley, O. Tickoo et al. , “Uncertainty as a form of transparency: Measuring, communicating, and using uncertainty,” in Proceedings of the 2021 AAAI/ACM Conference on AI, Ethics, and Society , 2021, pp. 401–413
2021
Cited alongside, same era.
E. Hüllermeier and W. Waegeman, “Aleatoric and epistemic uncertainty in machine learning: An introduction to concepts and methods,” Machine Learning , vol. 110, pp. 457–506, 2021
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2023
Closest in time.
D. Fried, A. Aghajanyan, J. Lin, S. Wang, E. Wallace, F. Shi, R. Zhong, S. Yih, L. Zettlemoyer, and M. Lewis, “Incoder: A generative model for code infilling and synthesis,” in The Eleventh International Conference on Learning Representations , 2023. [Online]. Available: https://openreview.net/forum?id=hQwb-lbM6EL
2023
Closest in time.
2023
Closest in time.
Q. Gu, “Llm-based code generation method for golang compiler testing,” in Proceedings of the 31st ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering , 2023, pp. 2201–2203
2023
Closest in time.
S. Kang, J. Yoon, and S. Yoo, “Large language models are few-shot testers: Exploring llm-based general bug reproduction,” in 2023 IEEE/ACM 45th International Conference on Software Engineering (ICSE) . IEEE, 2023, pp. 2312–2323
2023
Closest in time.
C. Lemieux, J. P. Inala, S. K. Lahiri, and S. Sen, “Codamosa: Escaping coverage plateaus in test generation with pre-trained large language models,” in International conference on software engineering (ICSE) , 2023
2023
Closest in time.
Z. Liu, C. Chen, J. Wang, X. Che, Y. Huang, J. Hu, and Q. Wang, “Fill in the blank: Context-aware automated text input generation for mobile gui testing,” in 2023 IEEE/ACM 45th International Conference on Software Engineering (ICSE) . IEEE, 2023, pp. 1355–1367
2023
Closest in time.
E. First, M. Rabe, T. Ringer, and Y. Brun, “Baldur: Whole-proof generation and repair with large language models,” in Proceedings of the 31st ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering , 2023, pp. 1229–1241
2023
Closest in time.
Y. Wei, C. S. Xia, and L. Zhang, “Copiloting the copilots: Fusing large language models with completion engines for automated program repair,” in Proceedings of the 31st ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering , 2023, pp. 172–184
2023
Closest in time.
M. Weiss and P. Tonella, “Uncertainty quantification for deep neural networks: An empirical comparison and usage guidelines,” Software Testing, Verification and Reliability , p. e1840, 2023
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
“Gpt 3.5,” https://platform.openai.com/docs/models/gpt-3-5 , 2023
2023
Closest in time.
C. Meister, T. Pimentel, G. Wiher, and R. Cotterell, “Locally typical sampling,” Transactions of the Association for Computational Linguistics , vol. 11, pp. 102–121, 2023
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
“all-mpnet-base-v2,” https://huggingface.co/sentence-transformers/all-mpnet-base-v2 , 2023
2023
Closest in time.
2023
Closest in time.
J. Ferrando, M. Sperber, H. Setiawan, D. Telaar, and S. Hasan, “Automating behavioral testing in machine translation,” in Proceedings of the Eighth Conference on Machine Translation , P. Koehn, B. Haddow, T. Kocmi, and C. Monz, Eds. Singapore: Association for Computational Linguistics, Dec. 2023, pp. 1014–1030. [Online]. Available: https://aclanthology.org/2023.wmt-1.97
2023
Closest in time.
C. S. Xia, Y. Wei, and L. Zhang, “Automated program repair in the era of large pre-trained language models,” in 2023 IEEE/ACM 45th International Conference on Software Engineering (ICSE) . IEEE, 2023, pp. 1482–1494
2023
Closest in time.
R. Tian, Y. Ye, Y. Qin, X. Cong, Y. Lin, Z. Liu, and M. Sun, “Debugbench: Evaluating debugging capability of large language models,” 2024
2024
Closest in time.
M. Geng, S. Wang, D. Dong, H. Wang, G. Li, Z. Jin, X. Mao, and X. Liao, “Large language models are few-shot summarizers: Multi-intent comment generation via in-context learning,” 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
“Gpt-4o,” https://openai.com/index/hello-gpt-4o/ , 2024
2024
Closest in time.
“Gpt-4o mini,” https://openai.com/index/gpt-4o-mini-advancing-cost-efficient-intelligence/ , 2024
2024
Closest in time.
J. Lee, S. Chen, A. Mordahl, C. Liu, W. Yang, and S. Wei, “Automated testing linguistic capabilities of nlp models,” ACM Transactions on Software Engineering and Methodology , 2024
2024
Closest in time.
2024
Closest in time.
R. Pan, A. R. Ibrahimzada, R. Krishna, D. Sankar, L. P. Wassi, M. Merler, B. Sobolev, R. Pavuluri, S. Sinha, and R. Jabbarvand, “Lost in translation: A study of bugs introduced by large language models while translating code,” in Proceedings of the IEEE/ACM 46th International Conference on Software Engineering , ser. ICSE ’24. New York, NY, USA: Association for Computing Machinery, 2024. [Online]. Available: https://doi-org.login.ezproxy.library.ualberta.ca/10.1145/3597503.3639226
2024
Closest in time.