Our new X account is live! Follow @wizwand_team for updates
WorkDL logo mark

Rethinking Graph Auto-Encoder Models for Attributed Graph Clustering

About

Most recent graph clustering methods have resorted to Graph Auto-Encoders (GAEs) to perform joint clustering and embedding learning. However, two critical issues have been overlooked. First, the accumulative error, inflicted by learning with noisy clustering assignments, degrades the effectiveness and robustness of the clustering model. This problem is called Feature Randomness. Second, reconstructing the adjacency matrix sets the model to learn irrelevant similarities for the clustering task. This problem is called Feature Drift. Interestingly, the theoretical relation between the aforementioned problems has not yet been investigated. We study these issues from two aspects: (1) there is a trade-off between Feature Randomness and Feature Drift when clustering and reconstruction are performed at the same level, and (2) the problem of Feature Drift is more pronounced for GAE models, compared with vanilla auto-encoder models, due to the graph convolutional operation and the graph decoding design. Motivated by these findings, we reformulate the GAE-based clustering methodology. Our solution is two-fold. First, we propose a sampling operator $\Xi$ that triggers a protection mechanism against the noisy clustering assignments. Second, we propose an operator $\Upsilon$ that triggers a correction mechanism against Feature Drift by gradually transforming the reconstructed graph into a clustering-oriented one. As principal advantages, our solution grants a considerable improvement in clustering effectiveness and robustness and can be easily tailored to existing GAE models.

Nairouz Mrabah, Mohamed Bouguessa, Mohamed Fawzi Touati, Riadh Ksantini• 2021

Related benchmarks

TaskDatasetResultRank
Node ClusteringCora
Accuracy76.7
115
Node ClusteringCiteseer
NMI45
110
ClusteringPubmed
Accuracy74
61
Graph ClusteringPubmed
Accuracy72.8
24
Graph ClusteringPubmed original
Accuracy74
12
ClusteringUSA Air-Traffic
Accuracy51.7
8
ClusteringEurope Air-Traffic
Accuracy57.4
8
ClusteringBrazil Air-Traffic
ACC74.1
8
Showing 8 of 8 rows

Other info

Code

Follow for update