Slice-Based Code Change Representation Learning

Fengyi Zhang,Yufei Zhao,Xin Peng,Bihuan Chen

doi:10.1109/saner56733.2023.00038

Abstract

Code changes are at the very core of software development and maintenance. Deep learning techniques have been used to build a model from a massive number of code changes to solve software engineering tasks, e.g., commit message generation and bug-fix commit identification. However, existing code change representation learning approaches represent code change as lexical tokens or syntactical AST (abstract syntax tree) paths, limiting the capability to learn semantics of code changes. Besides, they mostly do not consider noisy or tangled code change, hurting the accuracy of solved tasks. To address the above problems, we first propose a slice-based code change representation approach which considers data and control dependencies between changed code and unchanged code. Then, we propose a pre-trained sparse Transformer model, named CCS2VEC, to learn code change representations with three pre-training tasks. Our experiments by fine-tuning our pre-trained model on three downstream tasks have demonstrated the improvement of CCS2VEC over the state-of-the-art CC2VEC.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Slice-Based Code Change Representation Learning

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Automated Commit Intelligence by Pre-training
Shangqing Liu ... Wei Ma
ACM Transactions on Software Engineering and Methodology | VOL. -
Shangqing Liu, et. al.Shangqing Liu ... Wei Ma
01 Jul 2024
ACM Transactions on Software Engineering and Methodology | VOL. -

Measuring Task Similarity and Its Implication in Fine-Tuning Graph Neural Networks
Renhong Huang ... Jiarong Xu
Proceedings of the AAAI Conference on Artificial Intelligence | VOL. 38
Renhong Huang, et. al.Renhong Huang ... Jiarong Xu
24 Mar 2024
Proceedings of the AAAI Conference on Artificial Intelligence | VOL. 38

CodeEditor : Learning to Edit Source Code with Pre-trained Models
Jia Li ... Zhiyi Fu
ACM Transactions on Software Engineering and Methodology | VOL. 32
Jia Li, et. al.Jia Li ... Zhiyi Fu
30 Sep 2023
ACM Transactions on Software Engineering and Methodology | VOL. 32

PK-BERT: Knowledge Enhanced Pre-trained Models with Prompt for Few-Shot Learning
Han Ma ... Chan-Tong Lam
-
Han Ma, et. al.Han Ma ... Chan-Tong Lam
23 Nov 2022
23 Nov 2022

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Slice-Based Code Change Representation Learning

Abstract

Talk to us

Similar Papers