Fetching the paper…
Reading the bibliography…
Transformers are the current state-of-the-art of natural language processing in many domains and are using traction within software engineering research as well.
C. R. Blyth, “On simpson’s paradox and the sure-thing principle,” Journal of the American Statistical Association , vol. 67, no. 338, pp. 364–366, 1972
1972
Earlier work this paper cites.
T. D. Cook, D. T. Campbell, and A. Day, Quasi-experimentation: Design & analysis issues for field settings . Houghton Mifflin Boston, 1979, vol. 351
1979
Earlier work this paper cites.
S. Kiritchenko and S. Mohammad, “Examining gender and race bias in two hundred sentiment analysis systems,” in Proceedings of the Seventh Joint Conference on Lexical and Computational Semantics . New Orleans, Louisiana: Association for Computational Linguistics, Jun. 2018, pp. 43–53. [Online]. Available: https://aclanthology.org/S18-2005
2005
Earlier work this paper cites.
G. Antoniol, K. Ayari, M. Di Penta, F. Khomh, and Y.-G. Guéhéneuc, “Is it a bug or an enhancement?: A text-based approach to classify change requests,” in Proceedings of the 2008 Conference of the Center for Advanced Studies on Collaborative Research: Meeting of Minds , ser. CASCON ’08. New York, NY, USA: ACM, 2008, pp. 23:304–23:318. [Online]. Available: http://doi.acm.org/10.1145/1463788.1463819
2008
Earlier work this paper cites.
P. Runeson and M. Höst, “Guidelines for conducting and reporting case study research in software engineering,” Empirical Software Engineering , vol. 14, no. 2, pp. 131–164, 2009
2009
Earlier work this paper cites.
M. Thelwall, K. Buckley, G. Paltoglou, D. Cai, and A. Kappas, “Sentiment strength detection in short informal text,” Journal of the American Society for Information Science and Technology , vol. 61, no. 12, pp. 2544–2558, 2010. [Online]. Available: https://onlinelibrary.wiley.com/doi/abs/10.1002/asi.21416
2010
Earlier work this paper cites.
Internet Archive, “Internet Archive,” 2010. [Online]. Available: https://archive.org/index.php
2010
Earlier work this paper cites.
Google Cloud, “BigQuery - Google Cloud Platform.” 2011. [Online]. Available: https://cloud.google.com/bigquery/
2011
Earlier work this paper cites.
I. Grigorik, “The GitHub Archive,” 2012. [Online]. Available: https://www.gharchive.org/
2012
Earlier work this paper cites.
C. Wohlin, P. Runeson, M. Höst, M. C. Ohlsson, B. Regnell, and A. Wesslen, Experimentation in Software Engineering . Springer Publishing Company, Incorporated, 2012
2012
Earlier work this paper cites.
K. Herzig, S. Just, and A. Zeller, “It’s not a bug, it’s a feature: How misclassification impacts bug prediction,” in Proceedings of the International Conference on Software Engineering , ser. ICSE ’13. Piscataway, NJ, USA: IEEE Press, 2013, pp. 392–401. [Online]. Available: http://dl.acm.org/citation.cfm?id=2486788.2486840
2013
Earlier work this paper cites.
T. Mikolov, K. Chen, G. Corrado, and J. Dean, “Efficient estimation of word representations in vector space,” 2013
2013
Earlier work this paper cites.
S. Pyysalo, F. Ginter, H. Moen, T. Salakoski, and S. Ananiadou, “Distributional semantics resources for biomedical text processing,” in Proceedings of LBM 2013 , 2013, pp. 39–44
2013
Earlier work this paper cites.
Stack Exchange, “Stack Exchange Data Dump,” 2014. [Online]. Available: https://archive.org/details/stackexchange
2014
Earlier work this paper cites.
M. Ortu, G. Destefanis, B. Adams, A. Murgia, M. Marchesi, and R. Tonelli, “The jira repository dataset: Understanding social aspects of software development,” in Proceedings of the 11th International Conference on Predictive Models and Data Analytics in Software Engineering , ser. PROMISE ’15. New York, NY, USA: Association for Computing Machinery, 2015. [Online]. Available: https://doi.org/10.1145/2810146.2810147
2015
Earlier work this paper cites.
S. Nagel, “Cc-news,” 2016. [Online]. Available: https://commoncrawl.org/2016/10/news-dataset-available/
2016
Earlier work this paper cites.
S. Ghosh, P. Chakraborty, E. Cohn, J. S. Brownstein, and N. Ramakrishnan, “Characterizing diseases from unstructured text: A vocabulary driven word2vec approach,” 2016
2016
Earlier work this paper cites.
Stack Exchange, “Stack Overflow public dataset,” 2016. [Online]. Available: https://console.cloud.google.com/marketplace/product/stack-exchange/stack-overflow
2016
Earlier work this paper cites.
P. Rajpurkar, J. Zhang, K. Lopyrev, and P. Liang, “SQuAD: 100,000+ questions for machine comprehension of text,” in Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing . Austin, Texas: Association for Computational Linguistics, Nov. 2016, pp. 2383–2392. [Online]. Available: https://aclanthology.org/D16-1264
2016
Earlier work this paper cites.
W. Maalej, Z. Kurtanović, H. Nabil, and C. Stanik, “On the automatic classification of app reviews,” Requirements Engineering , vol. 21, no. 3, pp. 311–331, May 2016. [Online]. Available: https://doi.org/10.1007/s00766-016-0251-9
2016
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
T. Ahmed, A. Bosu, A. Iqbal, and S. Rahimi, “Senticr: A customized sentiment analysis tool for code review interactions,” in 2017 32nd IEEE/ACM International Conference on Automated Software Engineering (ASE) , 2017, pp. 106–111
2017
Earlier work this paper cites.
A. Benavoli, G. Corani, J. Demšar, and M. Zaffalon, “Time for a change: a tutorial for comparing multiple classifiers through bayesian analysis,” Journal of Machine Learning Research , vol. 18, no. 77, pp. 1–36, 2017. [Online]. Available: http://jmlr.org/papers/v18/16-305.html
2017
Earlier work this paper cites.
V. Efstathiou, C. Chatzilenas, and D. Spinellis, “Word embeddings for the software engineering domain,” in Proceedings of the 15th International Conference on Mining Software Repositories , ser. MSR ’18. New York, NY, USA: Association for Computing Machinery, 2018, p. 38–41. [Online]. Available: https://doi.org/10.1145/3196398.3196448
2018
Earlier work this paper cites.
M. R. Islam and M. F. Zibran, “Sentistrength-se: Exploiting domain specificity for improved sentiment analysis in software engineering text,” Journal of Systems and Software , vol. 145, pp. 125–146, 2018. [Online]. Available: https://www.sciencedirect.com/science/article/pii/S0164121218301675
2018
Cited alongside, same era.
F. Calefato, F. Lanubile, F. Maiorano, and N. Novielli, “Sentiment polarity detection for software development,” Empirical Softw. Engg. , vol. 23, no. 3, p. 1352–1382, jun 2018. [Online]. Available: https://doi.org/10.1007/s10664-017-9546-9
2018
Cited alongside, same era.
M. E. Peters, M. Neumann, M. Iyyer, M. Gardner, C. Clark, K. Lee, and L. Zettlemoyer, “Deep contextualized word representations,” in Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers) . New Orleans, Louisiana: Association for Computational Linguistics, Jun. 2018, pp. 2227–2237. [Online]. Available: https://aclanthology.org/N18-1202
2018
Cited alongside, same era.
J. Tabassum, M. Maddela, W. Xu, and A. Ritter, “Code and named entity recognition in StackOverflow,” in Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics . Online: Association for Computational Linguistics, Jul. 2020, pp. 4913–4926. [Online]. Available: https://www.aclweb.org/anthology/2020.acl-main.443
2020
Later among the works it cites.
S. Herbold, A. Trautsch, and F. Trautsch, “On the feasibility of automated prediction of bug and non-bug issues,” Empirical Software Engineering , vol. 25, no. 6, pp. 5333–5369, Sep. 2020. [Online]. Available: https://doi.org/10.1007/s10664-020-09885-w
2020
Later among the works it cites.
T. Zhang, B. Xu, F. Thung, S. A. Haryono, D. Lo, and L. Jiang, “Sentiment analysis for software engineering: How far can pre-trained transformer models go?” in 2020 IEEE International Conference on Software Maintenance and Evolution (ICSME) , 2020, pp. 70–80
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
P. Micikevicius, S. Narang, J. Alben, G. Diamos, E. Elsen, D. Garcia, B. Ginsburg, M. Houston, O. Kuchaiev, G. Venkatesh, and H. Wu, “Mixed precision training,” in International Conference on Learning Representations , 2018. [Online]. Available: https://openreview.net/forum?id=r1gs9JgRZ
2018
Cited alongside, same era.
N. Novielli, D. Girardi, and F. Lanubile, “A benchmark study on sentiment analysis for software engineering research,” in 2018 IEEE/ACM 15th International Conference on Mining Software Repositories (MSR) , 2018, pp. 364–375
2018
Cited alongside, same era.
B. Lin, F. Zampetti, G. Bavota, M. Di Penta, M. Lanza, and R. Oliveto, “Sentiment analysis for software engineering: How far can we go?” in Proceedings of the 40th International Conference on Software Engineering , ser. ICSE ’18. New York, NY, USA: Association for Computing Machinery, 2018, p. 94–104. [Online]. Available: https://doi.org/10.1145/3180155.3180195
2018
Cited alongside, same era.
B. Xu, A. Shirani, D. Lo, and M. A. Alipour, “Prediction of relatedness in stack overflow: Deep learning vs. svm: A reproducibility study,” in Proceedings of the 12th ACM/IEEE International Symposium on Empirical Software Engineering and Measurement , ser. ESEM ’18. New York, NY, USA: Association for Computing Machinery, 2018. [Online]. Available: https://doi.org/10.1145/3239235.3240503
2018
Cited alongside, same era.
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “BERT: Pre-training of deep bidirectional transformers for language understanding,” in Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) . Minneapolis, Minnesota: Association for Computational Linguistics, Jun. 2019, pp. 4171–4186. [Online]. Available: https://aclanthology.org/N19-1423
2019
Cited alongside, same era.
Z. Yang, Z. Dai, Y. Yang, J. Carbonell, R. Salakhutdinov, and Q. V. Le, XLNet: Generalized Autoregressive Pretraining for Language Understanding . Red Hook, NY, USA: Curran Associates Inc., 2019
2019
Cited alongside, same era.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, and I. Sutskever, “Language models are unsupervised multitask learners,” 2019
2019
Cited alongside, same era.
I. Beltagy, K. Lo, and A. Cohan, “Scibert: Pretrained language model for scientific text,” in EMNLP , 2019
2019
Cited alongside, same era.
J. Lee, W. Yoon, S. Kim, D. Kim, S. Kim, C. H. So, and J. Kang, “BioBERT: a pre-trained biomedical language representation model for biomedical text mining,” Bioinformatics , vol. 36, no. 4, pp. 1234–1240, 09 2019. [Online]. Available: https://doi.org/10.1093/bioinformatics/btz682
2019
Cited alongside, same era.
M. Lewis, Y. Liu, N. Goyal, M. Ghazvininejad, A. Mohamed, O. Levy, V. Stoyanov, and L. Zettlemoyer, “BART: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension,” in Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics . Online: Association for Computational Linguistics, Jul. 2020, pp. 7871–7880. [Online]. Available: https://aclanthology.org/2020.acl-main.703
2020
Later among the works it cites.
Z. Lan, M. Chen, S. Goodman, K. Gimpel, P. Sharma, and R. Soricut, “Albert: A lite bert for self-supervised learning of language representations,” in International Conference on Learning Representations , 2020. [Online]. Available: https://openreview.net/forum?id=H1eA7AEtvS
2020
Later among the works it cites.
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, S. Agarwal, A. Herbert-Voss, G. Krueger, T. Henighan, R. Child, A. Ramesh, D. M. Ziegler, J. Wu, C. Winter, C. Hesse, M. Chen, E. Sigler, M. Litwin, S. Gray, B. Chess, J. Clark, C. Berner, S. McCandlish, A. Radford, I. Sutskever, and D. Amodei, “Language models are few-shot learners,” 2020
2020
Later among the works it cites.
E. Biswas, M. E. Karabulut, L. Pollock, and K. Vijay-Shanker, “Achieving reliable sentiment analysis in the software engineering domain using bert,” in 2020 IEEE International Conference on Software Maintenance and Evolution (ICSME) , 2020, pp. 162–173
2020
Later among the works it cites.
Z. Feng, D. Guo, D. Tang, N. Duan, X. Feng, M. Gong, L. Shou, B. Qin, T. Liu, D. Jiang, and M. Zhou, “Codebert: A pre-trained model for programming and natural languages,” 2020
2020
Later among the works it cites.
K. Clark, M.-T. Luong, Q. V. Le, and C. D. Manning, “Electra: Pre-training text encoders as discriminators rather than generators,” 2020
2020
Later among the works it cites.
Y. You, J. Li, S. Reddi, J. Hseu, S. Kumar, S. Bhojanapalli, X. Song, J. Demmel, K. Keutzer, and C.-J. Hsieh, “Large batch optimization for deep learning: Training bert in 76 minutes,” in International Conference on Learning Representations , 2020. [Online]. Available: https://openreview.net/forum?id=Syx4wnEtvH
2020
Later among the works it cites.
S. Herbold, A. Trautsch, and F. Trautsch, “Issues with szz: An empirical assessment of the state of practice of defect prediction data collection,” 2020
2020
Later among the works it cites.
N. Novielli, F. Calefato, D. Dongiovanni, D. Girardi, and F. Lanubile, Can We Use SE-Specific Sentiment Analysis Tools in a Cross-Platform Setting? New York, NY, USA: Association for Computing Machinery, 2020, p. 158–168. [Online]. Available: https://doi.org/10.1145/3379597.3387446
2020
Later among the works it cites.
M. Zaheer, G. Guruganesh, K. A. Dubey, J. Ainslie, C. Alberti, S. Ontanon, P. Pham, A. Ravula, Q. Wang, L. Yang, and A. Ahmed, “Big bird: Transformers for longer sequences,” in Advances in Neural Information Processing Systems , H. Larochelle, M. Ranzato, R. Hadsell, M. F. Balcan, and H. Lin, Eds., vol. 33. Curran Associates, Inc., 2020, pp. 17 283–17 297. [Online]. Available: https://proceedings.neurips.cc/paper/2020/file/c8512d142a2d849725f31a9a7a361ab9-Paper.pdf
2020
Later among the works it cites.
Z. Feng, D. Guo, D. Tang, N. Duan, X. Feng, M. Gong, L. Shou, B. Qin, T. Liu, D. Jiang, and M. Zhou, “CodeBERT: A pre-trained model for programming and natural languages,” in Findings of the Association for Computational Linguistics: EMNLP 2020 . Online: Association for Computational Linguistics, Nov. 2020, pp. 1536–1547. [Online]. Available: https://aclanthology.org/2020.findings-emnlp.139
2020
Later among the works it cites.
A. Trautsch, J. Erbel, S. Herbold, and J. Grabowski, “On the differences between quality increasing and other changes in open source java projects,” 2021
2021
Closest in time.
W. Fedus, B. Zoph, and N. Shazeer, “Switch transformers: Scaling to trillion parameter models with simple and efficient sparsity,” 2021
2021
Closest in time.
L. Tunstall, L. von Werra, and T. Wolf, Natural Language Processing with Transformers . O’Reilly Media, Inc., 2021 (early access)
2021
Closest in time.
N. Novielli, F. Calefato, F. Lanubile, and A. Serebrenik, “Assessment of off-the-shelf SE-specific sentiment analysis tools: An extended replication study,” Empirical Software Engineering , vol. 26, no. 4, Jun. 2021. [Online]. Available: https://doi.org/10.1007/s10664-021-09960-w
2021
Closest in time.
G. Uddin and F. Khomh, “Automatic mining of opinions expressed about apis in stack overflow,” IEEE Transactions on Software Engineering , vol. 47, no. 3, pp. 522–559, 2021
2021
Closest in time.
“Ethnologue entry for english, 24th edition,” 2021. [Online]. Available: https://www.ethnologue.com/language/eng
2021
Closest in time.
“Ethnologue, 24th edition,” 2021. [Online]. Available: https://www.ethnologue.com/about/language-info
2021
Closest in time.
E. M. Bender, T. Gebru, A. McMillan-Major, and S. Shmitchell, “On the dangers of stochastic parrots: Can language models be too big?” in Proceedings of the 2021 ACM Conference on Fairness, Accountability, and Transparency , ser. FAccT ’21. New York, NY, USA: Association for Computing Machinery, 2021, p. 610–623. [Online]. Available: https://doi.org/10.1145/3442188.3445922
2021
Closest in time.
P. Icha, T. Lauf, and G. Kuhs, “Entwicklung der spezifischen Kohlendioxid-Emissionen des deutschen Strommix in den Jahren 1990 - 2019,” 2021
2021
Closest in time.
V. J. Hellendoorn and A. A. Sawant, “The growing cost of deep learning for source code,” Commun. ACM , vol. 65, no. 1, p. 31–33, dec 2021. [Online]. Available: https://doi.org/10.1145/3501261
2021
Closest in time.