Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

João Mattos

Attribute-Enhanced Similarity Ranking for Sparse Link Prediction

Nov 29, 2024

João Mattos, Zexi Huang, Mert Kosan, Ambuj Singh, Arlei Silva

Figure 1 for Attribute-Enhanced Similarity Ranking for Sparse Link Prediction

Figure 2 for Attribute-Enhanced Similarity Ranking for Sparse Link Prediction

Figure 3 for Attribute-Enhanced Similarity Ranking for Sparse Link Prediction

Figure 4 for Attribute-Enhanced Similarity Ranking for Sparse Link Prediction

Abstract:Link prediction is a fundamental problem in graph data. In its most realistic setting, the problem consists of predicting missing or future links between random pairs of nodes from the set of disconnected pairs. Graph Neural Networks (GNNs) have become the predominant framework for link prediction. GNN-based methods treat link prediction as a binary classification problem and handle the extreme class imbalance -- real graphs are very sparse -- by sampling (uniformly at random) a balanced number of disconnected pairs not only for training but also for evaluation. However, we show that the reported performance of GNNs for link prediction in the balanced setting does not translate to the more realistic imbalanced setting and that simpler topology-based approaches are often better at handling sparsity. These findings motivate Gelato, a similarity-based link-prediction method that applies (1) graph learning based on node attributes to enhance a topological heuristic, (2) a ranking loss for addressing class imbalance, and (3) a negative sampling scheme that efficiently selects hard training pairs via graph partitioning. Experiments show that Gelato outperforms existing GNN-based alternatives.

* To appear at the 31st SIGKDD Conference on Knowledge Discovery and Data Mining - Research Track (August 2024 Deadline)

Via

Access Paper or Ask Questions