arrow
Return

Layer-wise contrastive learning BERT for sentence representation of GitHub

delete2025-09-15
delete0
PRE
AI
D
Daoquan Chen
W
Wei Zhang
S
Shengyu Lu
林元国 cover
林元国 (Yuanguo Lin)
G
Gu, Xinyu
X
Xiuze Zhou *
DOI:10.1016/j.neucom.2025.131504delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Every day, on GitHub, end-users submit a large number of issues that must be addressed to ensure the success of software projects. Though BERT-based pre-trained language models achieve high performance on many downstream tasks, the sentence representation of [CLS] from the top layer of BERT has a limited ability to capture the semantic meaning of sentences. GitHub issue reports often include code snippets and user-generated terms not found in standard vocabularies. Therefore, the classification predictions of BERT are affected. To generate better sentence semantic representations of BERT for GitHub, we propose a layer-wise Contrastive Learning BERT (CLBERT), which uses contrastive learning to enhance the representation ability by contrasting the layer-by-layer representation. Further, to obtain as comprehensive information as possible, representations of each layer are extracted and learned by an attention mechanism as the final classification features. Finally, experiments conducted on two GitHub data sets show that our proposed model significantly improves classification performance.

Journal

Neurocomputing cover
Neurocomputing
IF:
6.5
Papers:
2.5W
Citations:
6.5W

Organization

H
Harbin Engineering University
Scholars:
1.9W
Papers: 1.3W
Citations: 1.3W
J
Jimei University
Scholars:
5.0K
Papers: 3.3K
Citations: 4.8K
S
Shanghai Artificial Intelligence Laboratory
Scholars:
472
Papers: 260
Citations: 765
X
xiamen university
Scholars:
5.8W
Papers: 3.8W
Citations: 67
researcher View more organizations