Fetching the paper…
Reading the bibliography…
Code search with natural language helps us reuse existing code snippets.
Étude comparative de la distribution florale dans une portion des alpes et des jura
Paul Jaccard. 1901 · 1901
Earlier work this paper cites.
Relevance weighting of search terms
Stephen E Robertson and K Sparck Jones. 1976 · 1976
Earlier work this paper cites.
Latent semantic indexing: An overview
Barbara Rosario. 2000 · 2000
Earlier work this paper cites.
Introduction to information retrieval , volume 39
Hinrich Schütze, Christopher D Manning, and Prabhakar Raghavan. 2008 · 2008
Earlier work this paper cites.
The probabilistic relevance framework: BM25 and beyond
Stephen Robertson and Hugo Zaragoza. 2009 · 2009
Earlier work this paper cites.
Improving source code search with natural language phrasal representations of method signatures
Emily Hill, Lori Pollock, and K Vijay-Shanker. 2011 · 2011
Earlier work this paper cites.
Codehow: Effective code search based on api understanding and extended boolean model (e)
Fei Lv, Hongyu Zhang, Jian-guang Lou, Shaowei Wang, Dongmei Zhang, and Jianjun Zhao. 2015 · 2015
Earlier work this paper cites.
Query expansion based on crowd knowledge for code search
Liming Nie, He Jiang, Zhilei Ren, Zeyi Sun, and Xiaochen Li. 2016 · 2016
Earlier work this paper cites.
A search log mining based query expansion technique to improve effectiveness in code search
Abdus Satter and Kazi Sakib. 2016 · 2016
Earlier work this paper cites.
Neural machine translation of rare words with subword units
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Earlier work this paper cites.
Combining word2vec with revised vector space model for better code retrieval
Thanh Van Nguyen, Anh Tuan Nguyen, Hung Dang Phan, Trong Duc Nguyen, and Tien N Nguyen. 2017 · 2017
Earlier work this paper cites.
Iecs: Intent-enforced code search via extended boolean model
Yangrui Yang and Qing Huang. 2017 · 2017
Earlier work this paper cites.
Deep code search
Xiaodong Gu, Hongyu Zhang, and Sunghun Kim. 2018 · 2018
Earlier work this paper cites.
Retrieval on source code: a neural code search
Saksham Sachdev, Hongyu Li, Sifei Luan, Seohyun Kim, Koushik Sen, and Satish Chandra. 2018 · 2018
Earlier work this paper cites.
Capturing source code semantics via tree-based convolution over api-enhanced ast
Long Chen, Wei Ye, and Shikun Zhang. 2019 · 2019
Earlier work this paper cites.
Generating long sequences with sparse transformers
Rewon Child, Scott Gray, Alec Radford, and Ilya Sutskever. 2019 · 2019
Earlier work this paper cites.
Adaptively sparse transformers
Gonçalo M Correia, Vlad Niculae, and André FT Martins. 2019 · 2019
Earlier work this paper cites.
Transformer-xl: Attentive language models beyond a fixed-length context
Zihang Dai, Zhilin Yang, Yiming Yang, Jaime G Carbonell, Quoc Le, and Ruslan Salakhutdinov. 2019 · 2019
Earlier work this paper cites.
BERT: pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
Star-transformer
Qipeng Guo, Xipeng Qiu, Pengfei Liu, Yunfan Shao, Xiangyang Xue, and Zheng Zhang. 2019 · 2019
Earlier work this paper cites.
Global relational models of source code
Vincent J Hellendoorn, Charles Sutton, Rishabh Singh, Petros Maniatis, and David Bieber. 2019 · 2019
Cited alongside, same era.
Codesearchnet challenge: Evaluating the state of semantic code search
Hamel Husain, Ho-Hsiang Wu, Tiferet Gazit, Miltiadis Allamanis, and Marc Brockschmidt. 2019 · 2019
Cited alongside, same era.
Reformer: The efficient transformer
Nikita Kitaev, Lukasz Kaiser, and Anselm Levskaya. 2019 · 2019
Cited alongside, same era.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 2019
Cited alongside, same era.
Multi-modal attention network learning for semantic source code retrieval
Yao Wan, Jingdong Shu, Yulei Sui, Guandong Xu, Zhou Zhao, Jian Wu, and Philip S. Yu. 2019 · 2019
Cited alongside, same era.
Multi-passage bert: A globally normalized bert model for open-domain question answering
Cosqa: 20,000+ web queries for code search and question answering
Junjie Huang, Duyu Tang, Linjun Shou, Ming Gong, Ke Xu, Daxin Jiang, Ming Zhou, and Nan Duan. 2021 · 2021
Later among the works it cites.
Code prediction by feeding trees to transformers
Seohyun Kim, Jinman Zhao, Yuchi Tian, and Satish Chandra. 2021 · 2021
Later among the works it cites.
Multi-modal multi-instance learning for retinal disease recognition
Xirong Li, Yang Zhou, Jie Wang, Hailan Lin, Jianchun Zhao, Dayong Ding, Weihong Yu, and Youxin Chen. 2021 · 2021
Later among the works it cites.
Integrating tree path in transformer for code representation
Han Peng, Ge Li, Wenhan Wang, Yunfei Zhao, and Zhi Jin. 2021 · 2021
Later among the works it cites.
Efficient content-based sparse attention with routing transformers
Aurko Roy, Mohammad Saffar, Ashish Vaswani, and David Grangier. 2021 · 2021
Later among the works it cites.
Cast: Enhancing code summarization with hierarchical splitting and reconstruction of abstract syntax trees
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhiguo Wang, Patrick Ng, Xiaofei Ma, Ramesh Nallapati, and Bing Xiang. 2019 · 2019
Cited alongside, same era.
Bp-transformer: Modelling long-range context via binary partitioning
Zihao Ye, Qipeng Guo, Quan Gan, Xipeng Qiu, and Zheng Zhang. 2019 · 2019
Cited alongside, same era.
ETC: encoding long and structured inputs in transformers
Joshua Ainslie, Santiago Ontañón, Chris Alberti, Vaclav Cvicek, Zachary Fisher, Philip Pham, Anirudh Ravula, Sumit Sanghai, Qifan Wang, and Li Yang. 2020 · 2020
Cited alongside, same era.
Structural language models of code
Uri Alon, Roy Sadaka, Omer Levy, and Eran Yahav. 2020 · 2020
Cited alongside, same era.
Longformer: The long-document transformer
Iz Beltagy, Matthew E Peters, and Arman Cohan. 2020 · 2020
Cited alongside, same era.
Codebert: A pre-trained model for programming and natural languages
Zhangyin Feng, Daya Guo, Duyu Tang, Nan Duan, Xiaocheng Feng, Ming Gong, Linjun Shou, Bing Qin, Ting Liu, Daxin Jiang, and Ming Zhou. 2020 · 2020
Cited alongside, same era.
Long document ranking with query-directed sparse transformer
Jyun-Yu Jiang, Chenyan Xiong, Chia-Jung Lee, and Wei Wang. 2020 · 2020
Cited alongside, same era.
Ensheng Shi, Yanlin Wang, Lun Du, Hongyu Zhang, Shi Han, Dongmei Zhang, and Hongbin Sun. 2021 · 2021
Later among the works it cites.
Cross-domain deep code search with few-shot meta learning
Yitian Chai, Hongyu Zhang, Beijun Shen, and Xiaodong Gu. 2022 · 2022
Closest in time.
Long document re-ranking with modular re-ranker
Luyu Gao and Jamie Callan. 2022 · 2022
Closest in time.
Heat: Hyperedge attention networks
Dobrik Georgiev, Marc Brockschmidt, and Miltiadis Allamanis. 2022 · 2022
Closest in time.
Accelerating code search with deep hashing and code classification
Wenchao Gu, Yanlin Wang, Lun Du, Hongyu Zhang, Shi Han, Dongmei Zhang, and Michael Lyu. 2022 · 2022
Closest in time.
Unixcoder: Unified cross-modal pre-training for code representation
Daya Guo, Shuai Lu, Nan Duan, Yanlin Wang, Ming Zhou, and Jian Yin. 2022 · 2022
Closest in time.
Lightweight attentional feature fusion: A new baseline for text-to-video retrieval
Fan Hu, Aozhu Chen, Ziyue Wang, Fangming Zhou, Jianfeng Dong, and Xirong Li. 2022 · 2022
Closest in time.
Code search based on context-aware code translation
Weisong Sun, Chunrong Fang, Yuchen Chen, Guanhong Tao, Tingxu Han, and Quanjun Zhang. 2022 · 2022
Closest in time.
Pre-training code representation with semantic flow graph for effective bug localization
Yali Du and Zhongxing Yu. 2023 · 2023
Closest in time.
Jina embeddings 2: 8192-token general-purpose text embeddings for long documents
Michael Günther, Jackmin Ong, Isabelle Mohr, Alaeddine Abdessalem, Tanguy Abel, Mohammad Kalim Akram, Susana Guzman, Georgios Mastrapas, Saba Sturua, Bo Wang, et al. 2023 · 2023
Closest in time.
Longcoder: A long-range pre-trained language model for code completion
Daya Guo, Canwen Xu, Nan Duan, Jian Yin, and Julian J. McAuley. 2023 · 2023
Closest in time.
Revisiting code search in a two-stage paradigm
Fan Hu, Yanlin Wang, Lun Du, Xirong Li, Hongyu Zhang, Shi Han, and Dongmei Zhang. 2023 · 2023
Closest in time.
Capturing the long-distance dependency in the control flow graph via structural-guided attention for bug localization
Y Ma, Yali Du, and Ming Li. 2023 · 2023
Closest in time.
Contextualized medication event extraction with striding ner and multi-turn qa
Tomoki Tsujimura, Koshi Yamada, Ryuki Ida, Makoto Miwa, and Yutaka Sasaki. 2023 · 2023
Closest in time.