Computer Science > Software Engineering

arXiv:2208.08014 (cs)

[Submitted on 17 Aug 2022 (v1), last revised 31 Aug 2022 (this version, v2)]

Title:AUGER: Automatically Generating Review Comments with Pre-training Models

Authors:Lingwei Li, Li Yang, Huaxi Jiang, Jun Yan, Tiejian Luo, Zihan Hua, Geng Liang, Chun Zuo

View PDF

Abstract:Code review is one of the best practices as a powerful safeguard for software quality. In practice, senior or highly skilled reviewers inspect source code and provide constructive comments, considering what authors may ignore, for example, some special cases. The collaborative validation between contributors results in code being highly qualified and less chance of bugs. However, since personal knowledge is limited and varies, the efficiency and effectiveness of code review practice are worthy of further improvement. In fact, it still takes a colossal and time-consuming effort to deliver useful review comments. This paper explores a synergy of multiple practical review comments to enhance code review and proposes AUGER (AUtomatically GEnerating Review comments): a review comments generator with pre-training models. We first collect empirical review data from 11 notable Java projects and construct a dataset of 10,882 code changes. By leveraging Text-to-Text Transfer Transformer (T5) models, the framework synthesizes valuable knowledge in the training stage and effectively outperforms baselines by 37.38% in ROUGE-L. 29% of our automatic review comments are considered useful according to prior studies. The inference generates just in 20 seconds and is also open to training further. Moreover, the performance also gets improved when thoroughly analyzed in case study.

Comments:	Accepted by ESEC/FSE-2022
Subjects:	Software Engineering (cs.SE)
Cite as:	arXiv:2208.08014 [cs.SE]
	(or arXiv:2208.08014v2 [cs.SE] for this version)
	https://doi.org/10.48550/arXiv.2208.08014

Submission history

From: Lingwei Li [view email]
[v1] Wed, 17 Aug 2022 01:38:01 UTC (1,466 KB)
[v2] Wed, 31 Aug 2022 13:30:59 UTC (1,881 KB)

Computer Science > Software Engineering

Title:AUGER: Automatically Generating Review Comments with Pre-training Models

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Software Engineering

Title:AUGER: Automatically Generating Review Comments with Pre-training Models

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators