首页|基于注意力机制的两阶段融合多视图图聚类

基于注意力机制的两阶段融合多视图图聚类

扫码查看
多视图图聚类旨在挖掘多视图图数据中蕴含的簇结构,近年来得到了研究者的广泛研究.然而,现有大多数方法在不同视图信息的融合过程中同等对待各个视图,未能根据视图质量分配相应权重,而且处理具有属性和图的数据时面临一定困难.该文提出了一种基于注意力机制的两阶段融合多视图图聚类算法.首先,应用图滤波器过滤高频噪声,各个视图获得更适用于下游聚类任务的节点平滑表示;其次,基于注意力机制融合各个视图特征滤波后的平滑表示,并为拓扑融合阶段提供初始化权重;然后,在拓扑融合阶段,将不同视图加权融合的Laplace矩阵与融合的特征表示输入编码器得到嵌入表示,并构造优化函数对权重和嵌入表示进行优化,可以为质量较好的视图分配较大权重,同时产生更加紧凑的嵌入表示;最后,通过对嵌入表示执行谱聚类得到最终的聚类结果.将该算法和已有的相关聚类算法在真实数据集上进行了实验分析.结果表明,相比已有算法,所提出的算法在处理多视图图数据方面更加有效.
Two-stage fusion multiview graph clustering based on the attention mechanism
[Objective]Multiview graph clustering aims to investigate the inherent cluster structures in multiview graph data and has received quite extensive research attention over recent years.However,there are differences in the final quality of different views,but existing methods treat all views equally during the fusion process without assigning the corresponding weights based on the received quality of the view.This may result in the loss of complementary information from multiple views and go on to ultimately affect the clustering quality.Additionally,the topological structure and attribute information of nodes in multiview graph data differ significantly in terms of content and form,making it somewhat challenging to integrate these two types of information effectively.To solve these problems,this paper proposes two-stage fusion multiview graph clustering based on an attention mechanism.[Methods]The algorithm can be divided into three stages:feature filtering based on graph filtering,feature fusion based on the attention mechanism,and topological fusion based on the attention mechanism.In the first stage,graph filters are applied to combine the attribute information with the topological structure of each view.In this process,a smoother embedding representation is achieved by filtering out high-frequency noise.In the second stage,the smooth representations of individual views are fused using attention mechanisms to obtain the consensus smooth representation,which incorporates information from all views.Additionally,a consensus Laplacian matrix is obtained by combining multiple views'Laplacian matrices using learnable weights.To obtain the final embedded representation,the consensus Laplacian matrix and consensus smooth representation are inputted into an encoder.Subsequently,the similarity matrix for the final embedded representation is computed.Training samples are selected from the similarity matrix,and the embedded representation and learnable weights of the Laplacian matrix are optimized iteratively to obtain a somewhat more compressed embedded representation.Finally,performing spectral clustering on the embedding representation yields the clustering results.The performance of the algorithm is evaluated using widely-used clustering evaluation metrics,including accuracy,normalized mutual information,an adjusted Rand index,and an F1-score,on three datasets:Association for Computing Machinery(ACM),Digital Bibliography & Library Project(DBLP),and Internet Movie Database(IMDB).[Results]1)The experimental results show that the proposed algorithm is more effective in handling multiview graph data,particularly for the ACM and DBLP datasets,compared to extant methods.However,it may not perform as well as LMEGC and MCGC on the IMDB dataset.2)Through the exploration of view quality using the proposed methods,the algorithm can learn weights specific to each view based on quality.3)Compared to the best-performing single view on each dataset(ACM,DBLP,and IMDB),the proposed algorithm achieves an average performance improvement of 2.4%,2.9%,and 2.1%,respectively,after fusing all views.4)Exploring the effect of the number of graph filter layers and the ratio of positive to negative node pairs on the performance of the algorithm,it was found that the best performance was achieved with somewhat small graph filter layers.The optimal ratio for positive and negative node pairs was around 0.01 and 0.5.[Conclusions]The algorithm combines attribute information with topological information through graph filtering to obtain smoother representations that are more suitable for clustering.The attention mechanisms can learn weights from both the topological and attribute information perspectives based on view quality.In this way,the representation could get the information from each view while avoiding the influence of poor-quality views.The proposed method in this paper achieves the expected results,greatly enhancing the clustering performance of the algorithm.

multiview learninggraph clusteringattention mechanismgraph learningembedded representation

赵兴旺、侯哲栋、姚凯旋、梁吉业

展开 >

山西大学计算机与信息技术学院,太原 030006

计算智能与中文信息处理教育部重点实验室(山西大学),太原 030006

多视图学习 图聚类 注意力机制 图学习 嵌入表示

国家自然科学基金面上项目国家自然科学基金面上项目国家自然科学基金区域联合创新基金重点项目

6207229362272285U21A20473

2024

清华大学学报(自然科学版)
清华大学

清华大学学报(自然科学版)

CSTPCD北大核心
影响因子:0.586
ISSN:1000-0054
年,卷(期):2024.64(1)
  • 2