Fetching the paper…
Reading the bibliography…
As innovation in deep learning continues, many engineers are incorporating Pre-Trained Models (PTMs) as components in computer systems.
Brown T, Mann B, Ryder N, Subbiah M, Kaplan JD, Dhariwal P, Neelakantan A, Shyam P, Sastry G, Askell A, et al. (2020a) Language models are few-shot learners. Advances in neural information processing systems 33:1877–1901
1901
Earlier work this paper cites.
Brown T, Mann B, Ryder N, Subbiah M, Kaplan JD, Dhariwal P, Neelakantan A, Shyam P, Sastry G, Askell A, et al. (2020b) Language models are few-shot learners. Advances in neural information processing systems 33:1877–1901
1901
Earlier work this paper cites.
Krueger CW (1992) Software reuse. ACM Computing Surveys (CSUR) 24(2):131–183
1992
Earlier work this paper cites.
Griss ML (1993) Software reuse: From library to factory. IBM systems journal 32(4):548–566
1993
Earlier work this paper cites.
Henninger S (1994) Using iterative refinement to find reusable software. IEEE software 11(5):48–59
1994
Earlier work this paper cites.
Seacord RC, Hissam SA, Wallnau KC (1998) Agora: A search engine for software components. IEEE Internet computing 2(6):62
1998
Earlier work this paper cites.
Guido van Rossum et al (2001) PEP 8 – Style Guide for Python Code. https://peps.python.org/pep-0008/
2001
Earlier work this paper cites.
Heineman GT, Councill WT (2001) Component-based software engineering. Putting the pieces together, addison-westley 5:1
2001
Earlier work this paper cites.
Szyperski C, Gruntz D, Murer S (2002) Component Software: Beyond Object-Oriented Programming, 2nd edn. Addison-Wesley
2002
Earlier work this paper cites.
González R, Van Der Meer K (2004) Standard metadata applied to software retrieval. Journal of Information Science 30(4):300–309
2004
Earlier work this paper cites.
Johnson RB, Onwuegbuzie AJ (2004) Mixed Methods Research: A Research Paradigm Whose Time Has Come. Educational Researcher
2004
Earlier work this paper cites.
2004
Earlier work this paper cites.
Sutter H, Alexandrescu A (2004) C++ coding standards: 101 rules, guidelines, and best practices. Pearson Education
2004
Earlier work this paper cites.
Frakes WB, Kang K (2005) Software reuse research: Status and future. IEEE transactions on Software Engineering 31(7):529–536
2005
Earlier work this paper cites.
Deissenboeck F, Pizka M (2006) Concise and consistent naming. Software Quality Journal 14:261–282
2006
Earlier work this paper cites.
Lau KK (2006) Software component models. In: Proceedings of the 28th international conference on Software engineering, pp 1081–1082
2006
Earlier work this paper cites.
Lawrie D, Morrell C, Feild H, Binkley D (2006) What’s in a name? a study of identifiers. In: 14th IEEE international conference on program comprehension (ICPC’06), IEEE, pp 3–12
2006
Earlier work this paper cites.
Lawrie D, Morrell C, Feild H, Binkley D (2007) Effective identifier names for comprehension and memory. Innovations in Systems and Software Engineering 3(4):303–318, DOI 10.1007/s11334-007-0031-2
2007
Earlier work this paper cites.
Boogerd C, Moonen L (2008) Assessing the value of coding standards: An empirical study. In: 2008 IEEE International conference on software maintenance, IEEE, pp 277–286
2008
Earlier work this paper cites.
Kitchenham BA, Pfleeger SL (2008) Personal opinion surveys. In: Guide to advanced empirical software engineering, Springer, pp 63–92
2008
Earlier work this paper cites.
Høst EW, Østvold BM (2009) Debugging method names. In: European Conference on Object-Oriented Programming, Springer, pp 294–317
2009
Earlier work this paper cites.
Ayodele TO (2010) Types of machine learning algorithms. New advances in machine learning 3(19-48):5–1
2010
Earlier work this paper cites.
Butler S, Wermelinger M, Yu Y, Sharp H (2010) Exploring the Influence of Identifier Names on Code Quality: An Empirical Study. In: 2010 14th European Conference on Software Maintenance and Reengineering, pp 156–165, DOI 10.1109/CSMR.2010.27
2010
Earlier work this paper cites.
Butler S, Wermelinger M, Yu Y, Sharp H (2011) Mining java class naming conventions. In: 2011 27th IEEE International Conference on Software Maintenance (ICSM), pp 93–102, DOI 10.1109/ICSM.2011.6080776
2011
Earlier work this paper cites.
Guest G, MacQueen KM, Namey EE (2011) Applied thematic analysis. sage publications
2011
Earlier work this paper cites.
Pradel M, Gross TR (2011) Detecting anomalies in the order of equally-typed method arguments. In: Proceedings of the 2011 International Symposium on Software Testing and Analysis, pp 232–242
2011
Earlier work this paper cites.
Wang S, Manning CD (2012) Baselines and bigrams: Simple, good sentiment and topic classification. In: Proceedings of the 50th Annual Meeting of the Association for Computational Linguistics (ACL), pp 90–94
2012
Earlier work this paper cites.
Wohlin C, Runeson P, Höst M, Ohlsson MC, Regnell B, Wesslén A, et al. (2012) Experimentation in software engineering, vol 236. Springer
2012
Earlier work this paper cites.
Bengio Y, Courville A, Vincent P (2013) Representation learning: A review and new perspectives. IEEE transactions on pattern analysis and machine intelligence 35(8):1798–1828
2013
Earlier work this paper cites.
Chung J, Gulcehre C, Cho K, Bengio Y (2014) Empirical evaluation of gated recurrent neural networks on sequence modeling. arXiv preprint arXiv:14123555
2014
Earlier work this paper cites.
Kim Y (2014) Convolutional neural networks for sentence classification. In: Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP), pp 1746–1751
2014
Earlier work this paper cites.
Kotonya G, Lee J (2014) Teaching reuse-driven software engineering through innovative role playing. In: Companion Proceedings of the 36th International Conference on Software Engineering, pp 276–282
2014
Earlier work this paper cites.
Olsson HH, Bosch J (2014) From opinions to data-driven software r&d: A multi-case study on how to close the’open loop’problem. In: 2014 40th EUROMICRO Conference on Software Engineering and Advanced Applications, IEEE, pp 9–16
2014
Earlier work this paper cites.
Allamanis M, Barr ET, Bird C, Sutton C (2015) Suggesting accurate method and class names. In: Proceedings of the 2015 10th Joint Meeting on Foundations of Software Engineering, ACM, Bergamo Italy, pp 38–49, DOI 10.1145/2786805.2786849
2015
Earlier work this paper cites.
LeCun Y, Bengio Y, Hinton G (2015) Deep learning. Nature 521(7553):436–444, DOI 10.1038/nature14539
2015
Earlier work this paper cites.
Sommerville I (2015) Software Engineering, 10th edn. Pearson Education Limited
2015
Earlier work this paper cites.
Zhang X, Zhao J, LeCun Y (2015) Character-level convolutional networks for text classification. In: Advances in Neural Information Processing Systems (NeurIPS)
2015
Earlier work this paper cites.
Zhu P, Zuo W, Zhang L, Hu Q, Shiu SC (2015) Unsupervised feature selection by regularized self-representation. Pattern Recognition 48(2):438–446
2015
Earlier work this paper cites.
Liu H, Liu Q, Staicu CA, Pradel M, Luo Y (2016) Nomen est omen: Exploring and exploiting similarities between argument and parameter names. In: Proceedings of the 38th International Conference on Software Engineering, pp 1063–1073
2016
Earlier work this paper cites.
Wittern E, Suter P, Rajagopalan S (2016) A look at the dynamics of the JavaScript package ecosystem. In: International Conference on Mining Software Repositories (MSR)
2016
Earlier work this paper cites.
Abdalkareem R, Nourry O, Wehaibi S, Mujahid S, Shihab E (2017) Why do developers use trivial packages? an empirical case study on npm. In: European Software Engineering Conference/Foundations of Software Engineering (ESEC/FSE)
2017
Earlier work this paper cites.
Hofmeister J, Siegmund J, Holt DV (2017a) Shorter identifier names take longer to comprehend. In: 2017 IEEE 24th International conference on software analysis, evolution and reengineering (SANER), IEEE, pp 217–227
2017
Earlier work this paper cites.
Hofmeister J, Siegmund J, Holt DV (2017b) Shorter identifier names take longer to comprehend. In: 2017 IEEE 24th International conference on software analysis, evolution and reengineering (SANER), IEEE, pp 217–227
2017
Earlier work this paper cites.
Sabor KK, Hamou-Lhadj A, Larsson A (2017) Durfex: a feature extraction technique for efficient detection of duplicate bug reports. In: 2017 IEEE international conference on software quality, reliability and security (QRS), IEEE, pp 240–250
2017
Earlier work this paper cites.
Terdchanakul P, Hata H, Phannachitta P, Matsumoto K (2017) Bug or not? bug report classification using n-gram idf. In: 2017 IEEE international conference on software maintenance and evolution (ICSME), IEEE, pp 534–538
2017
Earlier work this paper cites.
Vaswani A (2017) Attention is all you need. Advances in Neural Information Processing Systems
2017
Earlier work this paper cites.
Wu G, Cao Y, Chen W, Wei J, Zhong H, Huang T (2017) Appcheck: a crowdsourced testing service for android applications. In: 2017 IEEE International Conference on Web Services (ICWS), IEEE, pp 253–260
2017
Earlier work this paper cites.
Abiodun OI, Jantan A, Omolara AE, Dada KV, Mohamed NA, Arshad H (2018) State-of-the-art in artificial neural network applications: A survey. Heliyon 4(11)
2018
Earlier work this paper cites.
Habchi S, Blanc X, Rouvoy R (2018) On adopting linters to deal with performance concerns in android apps. In: Proceedings of the 33rd ACM/IEEE International Conference on Automated Software Engineering, pp 6–16
2018
Earlier work this paper cites.
Pradel M, Sen K (2018) Deepbugs: A learning approach to name-based bug detection. Proceedings of the ACM on Programming Languages 2(OOPSLA):1–25
2018
Earlier work this paper cites.
Schelter S, Biessmann F, Januschowski T, Salinas D, Seufert S, Szarvas G (2018) On Challenges in Machine Learning Model Management. Bulletin of the IEEE Computer Society Technical Committee on Data Engineering
2018
Earlier work this paper cites.
Tómasdóttir KF, Aniche M, Van Deursen A (2018) The adoption of javascript linters in practice: A case study on eslint. IEEE Transactions on Software Engineering 46(8):863–891
2018
Cited alongside, same era.
Violos J, Tserpes K, Varlamis I, Varvarigou T (2018) Text classification using the n-gram graph representation model over high frequency data streams. Frontiers in Applied Mathematics and Statistics 4:41
2018
Cited alongside, same era.
Decan A, Mens T, Grosjean P (2019) An empirical comparison of dependency network evolution in seven software packaging ecosystems. Empirical Software Engineering 24(1):381–416
2019
Cited alongside, same era.
Devlin J, Chang MW, Lee K, Toutanova K (2019) Bert: Pre-training of deep bidirectional transformers for language understanding. In: Proceedings of the 2019 conference of the North American chapter of the association for computational linguistics: human language technologies, volume 1 (long and short papers), pp 4171–4186
2019
2023
Closest in time.
Castaño J, Martínez-Fernández S, Franch X, Bogner J (2023) Analyzing the evolution and maintenance of ml models on hugging face. arXiv preprint arXiv:231113380
2023
Closest in time.
Checkmarx Research Team (2023) A new stealthier type of typosquatting attack spotted targeting npm. URL https://checkmarx.com/blog/a-new-stealthier-type-of-typosquatting-attack-spotted-targeting-npm/
2023
Closest in time.
Davis JC, Jajal P, Jiang W, Schorlemmer TR, Synovic N, Thiruvathukal GK (2023) Reusing deep learning models: Challenges and directions in software engineering. In: Proceedings of the IEEE John Vincent Atanasoff Symposium on Modern Computing (JVA’23)
2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Gao S, Chen C, Xing Z, Ma Y, Song W, Lin SW (2019) A Neural Model for Method Name Generation from Functional Description. In: 2019 IEEE 26th International Conference on Software Analysis, Evolution and Reengineering (SANER), pp 414–421, DOI 10.1109/SANER.2019.8667994
2019
Cited alongside, same era.
Gu T, Liu K, Dolan-Gavitt B, Garg S (2019) Badnets: Evaluating backdooring attacks on deep neural networks. IEEE Access 7:47230–47244
2019
Cited alongside, same era.
Géron A (2019) Hands-On Machine Learning with Scikit-Learn, Keras, and TensorFlow: Concepts, Tools, and Techniques to Build Intelligent Systems, 2nd edn. O’Reilly Media
2019
Cited alongside, same era.
Lu B, Yang J, Chen LY, Ren S (2019) Automating Deep Neural Network Model Selection for Edge Inference. In: 2019 IEEE First International Conference on Cognitive Machine Intelligence (CogMI), pp 184–193, DOI 10.1109/CogMI48466.2019.00035
2019
Cited alongside, same era.
Melegati J, Wang X, Abrahamsson P (2019) Hypotheses engineering: first essential steps of experiment-driven software development. In: 2019 IEEE/ACM Joint 4th International Workshop on Rapid Continuous Software Engineering and 1st International Workshop on Data-Driven Decisions, Experimentation and Evolution (RCoSE/DDrEE), IEEE, pp 16–19
2019
Cited alongside, same era.
Radford A, Wu J, Child R, Luan D, Amodei D, Sutskever I (2019) Language Models are Unsupervised Multitask Learners. OpenAI blog 1(8):9
2019
Cited alongside, same era.
Ren H, Xu B, Wang Y, Yi C, Huang C, Kou X, Xing T, Yang M, Tong J, Zhang Q (2019) Time-series anomaly detection service at microsoft. In: Proceedings of the 25th ACM SIGKDD international conference on knowledge discovery & data mining, pp 3009–3017
2019
Cited alongside, same era.
Soto-Valero C, Benelallam A, Harrand N, Barais O, Baudry B (2019) The emergence of software diversity in maven central. In: 2019 IEEE/ACM 16th International Conference on Mining Software Repositories (MSR), IEEE, pp 333–343
2019
Cited alongside, same era.
Garg G (2023) The power of hugging face ai. https://medium.com/@gargg/the-power-of-hugging-face-ai-4f6558ee0874
2023
Closest in time.
Geisler S, Li Y, Mankowitz DJ, Cemgil AT, Günnemann S, Paduraru C (2023) Transformers meet directed graphs. In: International Conference on Machine Learning (ICML’23), PMLR, pp 11144–11172
2023
Closest in time.
Gresta R, Durelli V, Cirilo E (2023) Naming practices in object-oriented programming: An empirical study. Journal of Software Engineering Research and Development pp 5–1
2023
Closest in time.
Gu Y, Ying L, Pu Y, Hu X, Chai H, Wang R, Gao X, Duan H (2023) Investigating package related security threats in software registries. In: 2023 IEEE Symposium on Security and Privacy (SP), IEEE, pp 1578–1595
2023
Closest in time.
Huang J, Lin B (2023) Cigar: Contrastive learning for github action recommendation. In: 2023 IEEE 23rd International Working Conference on Source Code Analysis and Manipulation (SCAM), IEEE, pp 61–71
2023
Closest in time.
Hugging Face (2023) Hugging face documentations. URL https://huggingface.co/docs
2023
Closest in time.
Infosecurity Magazine (2023) Malicious npm package uses new stealthier typosquatting attack. URL https://www.infosecurity-magazine.com/news/malicious-npm-package-uses/
2023
Closest in time.
Islam S, Elmekki H, Elsebai A, Bentahar J, Drawel N, Rjoub G, Pedrycz W (2023) A comprehensive survey on applications of transformers for deep learning tasks. Expert Systems with Applications p 122666
2023
Closest in time.
Kathikar A, Nair A, Lazarine B, Sachdeva A, Samtani S (2023) Assessing the Vulnerabilities of the Open-Source Artificial Intelligence (AI) Landscape: A Large-Scale Analysis of the Hugging Face Platform
2023
Closest in time.
Neupane S, Holmes G, Wyss E, Davidson D, De Carli L (2023) Beyond typosquatting: an in-depth look at package confusion. In: 32nd USENIX Security Symposium (USENIX Security 23), pp 3439–3456
2023
Closest in time.
NPM (2023) Package name guidelines. URL https://docs.npmjs.com/package-name-guidelines
2023
Closest in time.
ONNX (2023) Onnx model zoo. https://github.com/onnx/models
2023
Closest in time.
OpenBMB (2023) Issue #196 - project author team stay tuned: I found out that the llama3-v project is stealing a lot of academic work from minicpm-llama3-v 2.5. https://github.com/OpenBMB/MiniCPM-V/issues/196 , miniCPM-V. GitHub
2023
Closest in time.
Qi B, Sun H, Gao X, Zhang H, Li Z, Liu X (2023) Reusing deep neural network models through model re-engineering. In: 2023 IEEE/ACM 45th International Conference on Software Engineering (ICSE), IEEE, pp 983–994
2023
Closest in time.
ReversingLabs Threat Research Team (2023) r77 rootkit typosquatting: Npm threat research. URL https://www.reversinglabs.com/blog/r77-rootkit-typosquatting-npm-threat-research
2023
Closest in time.
Saieva A, Chakraborty S, Kaiser G (2023) On contrastive learning of semantic similarity forcode to code search. arXiv preprint arXiv:230503843
2023
Closest in time.
Taye MM (2023) Theoretical understanding of convolutional neural network: Concepts, architectures, applications, future directions. Computation 11(3):52
2023
Closest in time.
TensorFlow (2023) Tensorflow hub. URL https://www.tensorflow.org/hub
2023
Closest in time.
Verdecchia R, Engström E, Lago P, Runeson P, Song Q (2023) Threats to validity in software engineering research: A critical reflection. Information and Software Technology 164:107329
2023
Closest in time.
Wang C, Chen Z, Zhou M (2023a) AutoML from Software Engineering Perspective: Landscapes and Challenges. In: 2023 IEEE/ACM 20th International Conference on Mining Software Repositories (MSR), IEEE, Melbourne, Australia, pp 39–51, DOI 10.1109/MSR59073.2023.00019
2023
Closest in time.
Zamfirescu-Pereira J, Wong RY, Hartmann B, Yang Q (2023) Why johnny can’t prompt: how non-ai experts try (and fail) to design llm prompts. In: Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems, pp 1–21
2023
Closest in time.
Alpern R, Lazer I, Tzachor I, Hakim H, Weissbuch S, Feitelson DG (2024) Reproducing, extending, and analyzing naming experiments. arXiv
2024
Closest in time.
Di Sipio C, Rubei R, Di Rocco J, Di Ruscio D, Nguyen PT (2024) Automated categorization of pre-trained models for software engineering: A case study with a hugging face dataset. arXiv preprint arXiv:240513185
2024
Closest in time.
Hepworth I, Olive K, Dasgupta K, Le M, Lodato M, Maruseac M, Meiklejohn S, Chaudhuri S, Minkus T (2024) Securing the ai software supply chain. Technical report, Google
2024
Closest in time.
HiddenLayer (2024) Ai threat landscape report. https://21998286.fs1.hubspotusercontent-na1.net/hubfs/21998286/HiddenLayer%20AI%20Threat%20Landscape%20Report%202024.pdf , retrieved from HiddenLayer
2024
Closest in time.
Hugging Face, Inc (2024) Hugging face terms of service. URL https://huggingface.co/terms-of-service , accessed: 2024-08-18
2024
Closest in time.
Jajal P, Jiang W, Tewari A, Kocinare E, Woo J, Sarraf A, Lu YH, Thiruvathukal GK, Davis JC (2024) Interoperability in deep learning: A user survey and failure analysis of onnx model converters. the 33rd ACM SIGSOFT International Symposium on Software Testing and Analysis (ISSTA’24)
2024
Closest in time.
Jiang W, Yasmin J, Jones J, Synovic N, Kuo J, Tian Y, Thiruvathukal GK, , Davis JC (2024) Peatmoss: A dataset and initial analysis of pre-trained models in open-source software. In: Proceedings of the 21th Annual Conference on Mining Software Repositories (MSR’24)
2024
Closest in time.
Jones J, Jiang W, Synovic N, Thiruvathukal GK, Davis JC (2024) What do we know about hugging face? a systematic literature review and quantitative validation of qualitative claims. In: Proceedings of the 18th International Conference on Empirical Software Engineering and Measurement (ESEM)
2024
Closest in time.
Khalid H, Zimányi E (2024) Repairing raw metadata for metadata management. Information Systems 122:102344
2024
Closest in time.
Langford H, Shumailov I, Zhao Y, Mullins R, Papernot N (2024) Architectural neural backdoors from first principles. arXiv preprint arXiv:240206957
2024
Closest in time.
Li J, Tang T, Zhao WX, Nie JY, Wen JR (2024) Pre-trained language models for text generation: A survey. ACM Computing Surveys 56(9):1–39
2024
Closest in time.
ProtectAI (2024) Unveiling ai supply chain attacks on hugging face. URL https://protectai.com/threat-research/unveiling-ai-supply-chain-attacks-on-hugging-face
2024
Closest in time.
Sens Y, Knopp H, Peldszus S, Berger T (2024) A large-scale study of model integration in ml-enabled software systems. arXiv preprint arXiv:240806226
2024
Closest in time.
shersoni610, Li C (2024) Issue #1136: Naming llava models. https://github.com/haotian-liu/LLaVA/issues/1136
2024
Closest in time.
Tan X, Li T, Chen R, Liu F, Zhang L (2024) Challenges of using pre-trained models: the practitioners’ perspective. arXiv preprint arXiv:240414710
2024
Closest in time.
Taraghi M, Dorcelus G, Foundjem A, Tambon F, Khomh F (2024) Deep learning model reuse in the huggingface community: Challenges, benefit and trends. arXiv preprint arXiv:240113177
2024
Closest in time.
Zhang YK, Huang TJ, Ding YX, Zhan DC, Ye HJ (2024) Model spider: Learning to rank pre-trained models efficiently. Advances in Neural Information Processing Systems 36
2024
Closest in time.
DeepSeek AI (2024) Deepseek ai on hugging face. URL https://huggingface.co/deepseek-ai , accessed: 2025-04-14
2025
Closest in time.
Google (2024a) Google on hugging face. URL https://huggingface.co/google , accessed: 2025-04-14
2025
Closest in time.
Jiang W, Çakar B, Lysenko M, Davis JC (2025) Detecting active and stealthy typosquatting threats in package registries. arXiv preprint arXiv:250220528
2025
Closest in time.
Meta AI (2024) Meta ai on hugging face. URL https://huggingface.co/facebook , accessed: 2025-04-14
2025
Closest in time.
NVIDIA (2024) Inconsistent model config and architecture - issue #10006. https://github.com/NVIDIA/NeMo/issues/10006 , accessed: 2025-04-14
2025
Closest in time.
Patil PV, Jiang W, Peng H, Lugo D, Kalu KG, LeBlanc J, Smith L, Heo H, Aou N, Davis JC (2025) Recommending pre-trained models for iot devices. In: Proceedings of the International Workshop on Software Engineering Research in Practice (SERIP), to appear
2025
Closest in time.
Storey MA, Hoda R, Milani AMP, Baldassarre MT (2025) Guiding principles for using mixed methods research in software engineering. Empirical Software Engineering (EMSE)
2025
Closest in time.
Zhao J, Wang S, Zhao Y, Hou X, Wang K, Gao P, Zhang Y, Wei C, Wang H (2024) Models are codes: Towards measuring malicious code poisoning attacks on pre-trained model hubs. In: Proceedings of the 39th IEEE/ACM International Conference on Automated Software Engineering, pp 2087–2098
2098
Closest in time.