Rethinking Graph Auto-Encoder Models for Attributed Graph Clustering

Nairouz Mrabah, Mohamed Bouguessa, Mohamed Fawzi Touati, Riadh Ksantini

2021-07-19Graph Clustering Node Clustering Clustering

Abstract

Most recent graph clustering methods have resorted to Graph Auto-Encoders (GAEs) to perform joint clustering and embedding learning. However, two critical issues have been overlooked. First, the accumulative error, inflicted by learning with noisy clustering assignments, degrades the effectiveness and robustness of the clustering model. This problem is called Feature Randomness. Second, reconstructing the adjacency matrix sets the model to learn irrelevant similarities for the clustering task. This problem is called Feature Drift. Interestingly, the theoretical relation between the aforementioned problems has not yet been investigated. We study these issues from two aspects: (1) there is a trade-off between Feature Randomness and Feature Drift when clustering and reconstruction are performed at the same level, and (2) the problem of Feature Drift is more pronounced for GAE models, compared with vanilla auto-encoder models, due to the graph convolutional operation and the graph decoding design. Motivated by these findings, we reformulate the GAE-based clustering methodology. Our solution is two-fold. First, we propose a sampling operator $\Xi$ that triggers a protection mechanism against the noisy clustering assignments. Second, we propose an operator $\Upsilon$ that triggers a correction mechanism against Feature Drift by gradually transforming the reconstructed graph into a clustering-oriented one. As principal advantages, our solution grants a considerable improvement in clustering effectiveness and robustness and can be easily tailored to existing GAE models.

Results

Task	Dataset	Metric	Value	Model
Graph Clustering	Pubmed	ACC	74	R-GMM-VGAE
Graph Clustering	Pubmed	ARI	37.9	R-GMM-VGAE
Graph Clustering	Pubmed	NMI	33.4	R-GMM-VGAE
Graph Clustering	Pubmed	ACC	71.4	R-DGAE
Graph Clustering	Pubmed	ARI	34.6	R-DGAE
Graph Clustering	Pubmed	NMI	34.4	R-DGAE
Graph Clustering	Cora	ACC	76.7	R-GMM-VGAE
Graph Clustering	Cora	ARI	57.9	R-GMM-VGAE
Graph Clustering	Cora	NMI	57.3	R-GMM-VGAE
Graph Clustering	Cora	ACC	73.7	R-DGAE
Graph Clustering	Cora	ARI	54.1	R-DGAE
Graph Clustering	Cora	NMI	56	R-DGAE
Graph Clustering	Citeseer	ACC	70.5	R-DGAE
Graph Clustering	Citeseer	ARI	47.1	R-DGAE
Graph Clustering	Citeseer	NMI	45	R-DGAE
Graph Clustering	Citeseer	ACC	68.9	R-GMM-VGAE
Graph Clustering	Citeseer	ARI	43.9	R-GMM-VGAE
Graph Clustering	Citeseer	NMI	42	R-GMM-VGAE

Rethinking Graph Auto-Encoder Models for Attributed Graph Clustering

Abstract

Results

Related Papers

Rethinking Graph Auto-Encoder Models for Attributed Graph Clustering

Abstract

Results

Related Papers