Finding core labels for maximizing generalization of graph neural networks

Sichao Fu,Xueqi Ma,Yibing Zhan,Fanyu You,Qinmu Peng,Tongliang Liu,James Bailey,Danilo Mandic

doi:10.1016/j.neunet.2024.106635

Abstract

Graph neural networks (GNNs) have become a popular approach for semi-supervised graph representation learning. GNNs research has generally focused on improving methodological details, whereas less attention has been paid to exploring the importance of labeling the data. However, for semi-supervised learning, the quality of training data is vital. In this paper, we first introduce and elaborate on the problem of training data selection for GNNs. More specifically, focusing on node classification, we aim to select representative nodes from a graph used to train GNNs to achieve the best performance. To solve this problem, we are inspired by the popular lottery ticket hypothesis, typically used for sparse architectures, and we propose the following subset hypothesis for graph data: “There exists a core subset when selecting a fixed-size dataset from the dense training dataset, that can represent the properties of the dataset, and GNNs trained on this core subset can achieve a better graph representation”. Equipped with this subset hypothesis, we present an efficient algorithm to identify the core data in the graph for GNNs. Extensive experiments demonstrate that the selected data (as a training set) can obtain performance improvements across various datasets and GNNs architectures.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Finding core labels for maximizing generalization of graph neural networks

Abstract

Talk to us

Similar Papers

More From: Neural Networks

Lead the way for us

Similar Papers

Self-Paced Co-Training of Graph Neural Networks for Semi-Supervised Node Classification.
Maoguo Gong ... A K Qin
IEEE Transactions on Neural Networks and Learning Systems | VOL. 34
Maoguo Gong, et. al.Maoguo Gong ... A K Qin
01 Nov 2023
IEEE Transactions on Neural Networks and Learning Systems | VOL. 34

Violin: Virtual Overbridge Linking for Enhancing Semi-supervised Learning on Graphs with Limited Labels
Siyue Xie ... Da Sun Handason Tam
-
Siyue Xie, et. al.Siyue Xie ... Da Sun Handason Tam
01 Aug 2023
01 Aug 2023

GTC: GNN-Transformer co-contrastive learning for self-supervised heterogeneous graph representation
Yundong Sun ... Zhaoshuo Tian
Neural Networks | VOL. 181
Yundong Sun, et. al.Yundong Sun ... Zhaoshuo Tian
16 Aug 2024
Neural Networks | VOL. 181

Parallel and Distributed Graph Neural Networks: An In-Depth Concurrency Analysis.
Maciej Besta ... Torsten Hoefler
IEEE Transactions on Pattern Analysis and Machine Intelligence | VOL. 46
Maciej Besta, et. al.Maciej Besta ... Torsten Hoefler
01 May 2024
IEEE Transactions on Pattern Analysis and Machine Intelligence | VOL. 46

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Finding core labels for maximizing generalization of graph neural networks

Abstract

Talk to us

Similar Papers

More From: Neural Networks