2019

Neural Code Search Evaluation Dataset

Li, Hongyu, Kim, Seohyun, Chandra, Satish

Understand

There has been an increase of interest in code search using natural language.

  • Assessing the performance of such code search models can be difficult without a readily available evaluation suite.
  • In this paper, we present an evaluation dataset consisting of natural language query and code snippet pairs, with the hope that future work in this area can use this dataset as a common benchmark.
  • We also provide the results of two code search models ([1] and [6]) from recent work.

Built on

  • When deep learning met code search

    Original

    Jose Cambronero, Hongyu Li Seohyun Kim, Koushik Sen, and Satish Chandra · 1905

    Earlier work this paper cites.

  • Deep code search

    Xiaodong Gu, Hongyu Zhang, and Sunghun Kim · 2018

    Earlier work this paper cites.

Similar

Then

  • Retrieval on source code: a neural code search

    Saksham Sachdev, Hongyu Li, Sifei Luan, Seohyun Kim, Koushik Sen, and Satish Chandra · 2018

    Later among the works it cites.

Beyond the bibliography

alphaXiv searches the wider corpus for related work and actual follow-ups.

Open on alphaXiv

alphaXiv is searching for related work…