paper-with-me

홈 › Papers

Pitfalls in Link Prediction with Graph Neural Networks: Understanding the Impact of Target-link Inclusion & Better Practices

2023-06-01 · Jing Zhu, YuHang Zhou, Vassilis N. Ioannidis, Shengyi Qian, Wei Ai, Xiang Song, Danai Koutra

While Graph Neural Networks (GNNs) are remarkably successful in a variety of high-impact applications, we demonstrate that, in link prediction, the common practices of including the edges being predicted in the graph at training and/or test have outsized impact on the performance of low-degree nodes. We theoretically and empirically investigate how these practices impact node-level performance across different degrees. Specifically, we explore three issues that arise: (I1) overfitting; (I2) distribution shift; and (I3) implicit test leakage. The former two issues lead to poor generalizability to the test data, while the latter leads to overestimation of the model's performance and directly impacts the deployment of GNNs. To address these issues in a systematic way, we introduce an effective and efficient GNN training framework, SpotTarget, which leverages our insight on low-degree nodes: (1) at training time, it excludes a (training) edge to be predicted if it is incident to at least one low-degree node; and (2) at test time, it excludes all test edges to be predicted (thus, mimicking real scenarios of using GNNs, where the test data is not included in the graph). SpotTarget helps researchers and practitioners adhere to best practices for learning from graph data, which are frequently overlooked even by the most widely-used frameworks. Our experiments on various real-world datasets show that SpotTarget makes GNNs up to 15x more accurate in sparse graphs, and significantly improves their performance for low-degree nodes in dense graphs.

📄 PDF Abstract BibTeX arXiv:2306.00899

Code (0)

등록된 구현이 없습니다.

Tasks

Link PredictionNode Classification

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Evaluating Graph Neural Networks for Link Prediction: Current Pitfalls and New Benchmarking

2023-06-18 · NeurIPS 2023 11 · Juanhui Li, Harry Shomer, Haitao Mao, Shenglai Zeng 외

Link prediction attempts to predict whether an unseen edge exists based on only a portion of edges of a graph. A flurry of methods have been introduced in recent years that attempt to make use of graph neural networks (G…

BenchmarkingLink Prediction

Just Propagate: Unifying Matrix Factorization, Network Embedding, and LightGCN for Link Prediction

2024-10-26 · Haoxin Liu

Link prediction is a fundamental task in graph analysis. Despite the success of various graph-based machine learning models for link prediction, there lacks a general understanding of different models. In this paper, we …

Graph Neural NetworkLink PredictionNetwork EmbeddingPrediction

Pitfalls in Language Models for Code Intelligence: A Taxonomy and Survey

2023-10-27 · Xinyu She, Yue Liu, Yanjie Zhao, Yiling He 외

Modern language models (LMs) have been successfully employed in source code generation and understanding, leading to a significant increase in research focused on learning-based code intelligence, such as automated bug r…

Code Generation

Single-GPU GNN Systems: Traps and Pitfalls

2024-02-05 · Yidong Gong, Arnab Tarafder, Saima Afrin, Pradeep Kumar

The current graph neural network (GNN) systems have established a clear trend of not showing training accuracy results, and directly or indirectly relying on smaller datasets for evaluations majorly. Our in-depth analysi…

GPUGraph Neural Network

Natural Language Processing for Drug Discovery Knowledge Graphs: promises and pitfalls

2023-10-24 · J. Charles G. Jeynes, Tim James, Matthew Corney

Building and analysing knowledge graphs (KGs) to aid drug discovery is a topical area of research. A salient feature of KGs is their ability to combine many heterogeneous data sources in a format that facilitates discove…

Drug DiscoveryKnowledge Graphsnamed-entity-recognitionNamed Entity Recognition