Fetching the paper…
Reading the bibliography…
This paper tackles the challenge of teaching code semantics to Large Language Models (LLMs) for program analysis by incorporating code symmetries into the model architecture.
Two sides of the same coin: Exploiting the impact of identifiers in neural code comprehension
Gao, S., Gao, C., Wang, C., Sun, J., Lo, D., and Yu, Y · 1945
Earlier work this paper cites.
Permutations, matrices, and generalized young tableaux
Knuth, D · 1970
Earlier work this paper cites.
Algebraic graph theory
Biggs, N., Biggs, N. L., and Norman, B · 1993
Earlier work this paper cites.
Functional analysis and semi-groups , volume 31
Hille, E. and Phillips, R. S · 1996
Earlier work this paper cites.
Introduction to graph theory , volume 2
West, D. B. et al · 2001
Earlier work this paper cites.
A few billion lines of code later: using static analysis to find bugs in the real world
Bessey, A., Block, K., Chelf, B., Chou, A., Fulton, B., Hallem, S., Henri-Gros, C., Kamsky, A., McPeak, S., and Engler, D · 2010
Earlier work this paper cites.
Defects4j: A database of existing faults to enable controlled testing studies for java programs
Just, R., Jalali, D., and Ernst, M. D · 2014
Earlier work this paper cites.
A convolutional attention network for extreme summarization of source code
Allamanis, M., Peng, H., and Sutton, C · 2016
Earlier work this paper cites.
Group equivariant convolutional networks
Cohen, T. and Welling, M · 2016
Earlier work this paper cites.
Harnessing deep neural networks with logic rules
Hu, Z., Ma, X., Liu, Z., Hovy, E., and Xing, E · 2016
Earlier work this paper cites.
Learning to represent programs with graphs
Allamanis, M., Brockschmidt, M., and Khademi, M · 2017
Earlier work this paper cites.
Geometric deep learning: going beyond euclidean data
Bronstein, M. M., Bruna, J., LeCun, Y., Szlam, A., and Vandergheynst, P · 2017
Earlier work this paper cites.
Neural nets can learn function type signatures from binaries
Chua, Z. L., Shen, S., Saxena, P., and Liang, Z · 2017
Earlier work this paper cites.
Kim, Y., Denton, C., Hoang, L., and Rush, A. M · 2017
Earlier work this paper cites.
Neural network-based graph embedding for cross-platform binary code similarity detection
Xu, X., Liu, C., Feng, Q., Yin, H., Song, L., and Song, D · 2017
Earlier work this paper cites.
Hikari – an improvement over Obfuscator-LLVM
Zhang, N · 2017
Earlier work this paper cites.
code2seq: Generating sequences from structured representations of code
Alon, U., Brody, S., Levy, O., and Yahav, E · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M.-W., Lee, K., and Toutanova, K · 2018
Earlier work this paper cites.
Learning so (3) equivariant representations with spherical cnns
Esteves, C., Allen-Blanchette, C., Makadia, A., and Daniilidis, K · 2018
Earlier work this paper cites.
Structured neural summarization
Fernandes, P., Allamanis, M., and Brockschmidt, M · 2018
Earlier work this paper cites.
Towards a definition of disentangled representations
Higgins, I., Amos, D., Pfau, D., Racaniere, S., Matthey, L., Rezende, D., and Lerchner, A · 2018
Earlier work this paper cites.
Neural-symbolic vqa: Disentangling reasoning from vision and language understanding
Yi, K., Wu, J., Gan, C., Torralba, A., Kohli, P., and Tenenbaum, J · 2018
Earlier work this paper cites.
code2vec: Learning distributed representations of code
Alon, U., Zilberstein, M., Levy, O., and Yahav, E · 2019
Earlier work this paper cites.
Deep ensembles: A loss landscape perspective
Fort, S., Hu, H., and Lakshminarayanan, B · 2019
Earlier work this paper cites.
Permutation equivariant models for compositional generalization in language
Gordon, J., Lopez-Paz, D., Baroni, M., and Bouchacourt, D · 2019
Earlier work this paper cites.
DEEPVSA: Facilitating value-set analysis with deep learning for postmortem program analysis
Guo, W., Mu, D., Xing, X., Du, M., and Song, D · 2019
Earlier work this paper cites.
Global relational models of source code
Hellendoorn, V. J., Sutton, C., Singh, R., Maniatis, P., and Bieber, D · 2019
Earlier work this paper cites.
A mathematical view of attention models in deep learning
Ji, S., Xie, Y., and Gao, H · 2019
Cited alongside, same era.
Set transformer: A framework for attention-based permutation-invariant neural networks
Lee, J., Lee, Y., Kim, J., Kosiorek, A., Choi, S., and Teh, Y. W · 2019
Cited alongside, same era.
Fairseq: A fast, extensible toolkit for sequence modeling
Ott, M., Edunov, S., Baevski, A., Fan, A., Gross, S., Ng, N., Grangier, D., and Auli, M · 2019
Cited alongside, same era.
Pytorch: An imperative style, high-performance deep learning library
Paszke, A., Gross, S., Massa, F., Lerer, A., Bradbury, J., Chanan, G., Killeen, T., Lin, Z., Gimelshein, N., Antiga, L., et al · 2019
Cited alongside, same era.
Deepsphere: Efficient spherical convolutional neural network with healpix sampling for cosmological applications
Perraudin, N., Defferrard, M., Kacprzak, T., and Sgier, R · 2019
Cited alongside, same era.
Wang, Y., Wang, W., Joty, S., and Hoi, S. C · 2021
Later among the works it cites.
Do transformers really perform badly for graph representation?
Ying, C., Cai, T., Luo, S., Zheng, S., Ke, G., He, D., Shen, Y., and Liu, T.-Y · 2021
Later among the works it cites.
Dos and don’ts of machine learning in computer security
Arp, D., Quiring, E., Pendlebury, F., Warnecke, A., Pierazzi, F., Wressnegger, C., Cavallaro, L., and Rieck, K · 2022
Later among the works it cites.
Bieber, D., Goel, R., Zheng, D., Larochelle, H., and Tarlow, D · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bieber, D., Sutton, C., Larochelle, H., and Tarlow, D · 2020
Cited alongside, same era.
Lorentz group equivariant neural network for particle physics
Bogatskiy, A., Anderson, B., Offermann, J., Roussi, M., Miller, D., and Kondor, R · 2020
Cited alongside, same era.
Codebert: A pre-trained model for programming and natural languages
Feng, Z., Guo, D., Tang, D., Duan, N., Feng, X., Gong, M., Shou, L., Qin, B., Liu, T., Jiang, D., et al · 2020
Cited alongside, same era.
Graphcodebert: Pre-training code representations with data flow
Guo, D., Ren, S., Lu, S., Feng, Z., Tang, D., Liu, S., Zhou, L., Duan, N., Svyatkovskiy, A., Fu, S., et al · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., and Liu, P. J · 2020
Cited alongside, same era.
Semantic robustness of models of source code
Ramakrishnan, G., Henkel, J., Wang, Z., Albarghouthi, A., Jha, S., and Reps, T · 2020
Cited alongside, same era.
Group equivariant stand-alone self-attention for vision
Romero, D. W. and Cordonnier, J.-B · 2020
Cited alongside, same era.
Bundt, J., Davinroy, M., Agadakos, I., Oprea, A., and Robertson, W · 2022
Later among the works it cites.
Neural-symbolic learning and reasoning: A survey and interpretation
Garcez, A. d., Bader, S., Bowman, H., Lamb, L. C., de Penning, L., Illuminoo, B., Poon, H., and Zaverucha, C. G · 2022
Later among the works it cites.
Unixcoder: Unified cross-modal pre-training for code representation
Guo, D., Lu, S., Duan, N., Wang, Y., Zhou, M., and Yin, J · 2022
Later among the works it cites.
Semantic robustness of models of source code
Henke, J., Ramakrishnan, G., Wang, Z., Albarghouth, A., Jha, S., and Reps, T · 2022
Later among the works it cites.
BigCode is an open scientific collaboration working on the responsible development and use of large language models for code
HuggingFace and ServiceNow · 2022
Later among the works it cites.
Symlm: Predicting function names in stripped binaries via context-sensitive execution-aware code embeddings
Jin, X., Pei, K., Won, J. Y., and Lin, Z · 2022
Later among the works it cites.
How machine learning is solving the binary function similarity problem
Marcelli, A., Graziano, M., Ugarte-Pedrero, X., Fratantonio, Y., Mansouri, M., and Balzarotti, D · 2022
Later among the works it cites.
Trex: Learning execution semantics from micro-traces for binary similarity
Pei, K., Xuan, Z., Yang, J., Jana, S., and Ray, B · 2022
Later among the works it cites.
Graph neural networks for materials science and chemistry
Reiser, P., Neubert, M., Eberhard, A., Torresi, L., Zhou, C., Shao, C., Metni, H., van Hoesel, C., Schopmans, H., Sommer, T., et al · 2022
Later among the works it cites.
Recode: Robustness evaluation of code generation models
Wang, S., Li, Z., Qian, H., Yang, C., Wang, Z., Shang, M., Kumar, V., Tan, S., Ray, B., Bhatia, P., et al · 2022
Later among the works it cites.
Neural program repair with execution-based backpropagation
Ye, H., Martinez, M., and Monperrus, M · 2022
Later among the works it cites.
Equi-tuning: Group equivariant fine-tuning of pretrained models
Basu, S., Sattigeri, P., Ramamurthy, K. N., Chenthamarakshan, V., Varshney, K. R., Varshney, L. R., and Das, P · 2023
Closest in time.
Traced: Execution-aware pre-training for source code
Ding, Y., Steenhoek, B., Pei, K., Kaiser, G., Le, W., and Ray, B · 2023
Closest in time.
Scallop: A language for neurosymbolic programming
Li, Z., Huang, J., and Naik, M · 2023
Closest in time.
AI-Powered Fuzzing: Breaking the Bug Hunting Barrier
Liu, D., Metzman, J., and Chang, O · 2023
Closest in time.
Wizardcoder: Empowering code large language models with evol-instruct
Luo, Z., Xu, C., Zhao, P., Sun, Q., Geng, X., Hu, W., Tao, C., Ma, J., Lin, Q., and Jiang, D · 2023
Closest in time.
Large sequence models for software development activities
Maniatis, P. and Tarlow, D · 2023
Closest in time.
Code llama: Open foundation models for code
Roziere, B., Gehring, J., Gloeckle, F., Sootla, S., Gat, I., Tan, X. E., Adi, Y., Liu, J., Remez, T., Rapin, J., et al · 2023
Closest in time.
Lexecutor: Learning-guided execution
Souza, B. and Pradel, M · 2023
Closest in time.
Can large language models identify and reason about security vulnerabilities? not yet
Ullah, S., Han, M., Pujar, S., Pearce, H., Coskun, A., and Stringhini, G · 2023
Closest in time.
Pelican: Exploiting backdoors of naturally trained deep learning models in binary code analysis
Zhang, Z., Tao, G., Shen, G., An, S., Xu, Q., Liu, Y., Ye, Y., Wu, Y., and Zhang, X · 2023
Closest in time.
Discovering symmetry group structures via implicit orthogonality bias
Huh, D · 2024
Closest in time.
Grounding data science code generation with input-output specifications
Wen, Y., Yin, P., Shi, K., Michalewski, H., Chaudhuri, S., and Polozov, A · 2024
Closest in time.