首页> 外文OA文献 >A-Wardpβ: effective hierarchical clustering using the Minkowski metric and a fast k-means initialisation
【2h】

A-Wardpβ: effective hierarchical clustering using the Minkowski metric and a fast k-means initialisation

机译:A-Wardpβ:使用Minkowski度量和快速k-means初始化的有效分层聚类

代理获取
本网站仅为用户提供外文OA文献查询和代理获取服务,本网站没有原文。下单后我们将采用程序或人工为您竭诚获取高质量的原文,但由于OA文献来源多样且变更频繁,仍可能出现获取不到、文献不完整或与标题不符等情况,如果获取不到我们将提供退款服务。请知悉。

摘要

In this paper we make two novel contributions to hierarchical clustering. First, we introduce an anomalous pattern initialisation method for hierarchical clustering algorithms, called A-Ward, capable of substantially reducing the time they take to converge. This method generates an initial partition with a sufficiently large number of clusters. This allows the cluster merging process to start from this partition rather than from a trivial partition composed solely of singletons.ududOur second contribution is an extension of the Ward and Wardp algorithms to the situation where the feature weight exponent can differ from the exponent of the Minkowski distance. This new method, called A-Wardpβ, is able to generate a much wider variety of clustering solutions. We also demonstrate that its parameters can be estimated reasonably well by using a cluster validity index.ududWe perform numerous experiments using data sets with two types of noise, insertion of noise features and blurring within-cluster values of some features. These experiments allow us to conclude: (i) our anomalous pattern initialisation method does indeed reduce the time a hierarchical clustering algorithm takes to complete, without negatively impacting its cluster recovery ability; (ii) A-Wardpβ provides better cluster recovery than both Ward and Wardp.
机译:在本文中,我们对层次聚类做出了两个新颖的贡献。首先,我们为分层聚类算法引入了一种称为A-Ward的异常模式初始化方法,该方法能够大大减少收敛所需的时间。此方法生成具有足够大量群集的初始分区。这使群集合并过程可以从该分区开始,而不是从仅由单例组成的琐碎分区开始。 ud ud我们的第二个贡献是Ward和Wardp算法的扩展,以解决特征权重指数可能不同于指数的情况Minkowski距离的这种称为A-Wardpβ的新方法能够生成种类繁多的聚类解决方案。我们还证明了通过使用聚类有效性指标可以合理地估计其参数。 ud ud我们使用具有两种类型的噪声,噪声特征的插入和某些特征的簇内值模糊的数据集进行了许多实验。这些实验使我们可以得出以下结论:(i)我们的异常模式初始化方法确实确实减少了分层聚类算法完成的时间,而没有负面影响其聚类恢复能力; (ii)与Ward和Wardp相比,A-Wardpβ提供更好的群集恢复。

著录项

相似文献

  • 外文文献
  • 中文文献
  • 专利
代理获取

客服邮箱:kefu@zhangqiaokeyan.com

京公网安备:11010802029741号 ICP备案号:京ICP备15016152号-6 六维联合信息科技 (北京) 有限公司©版权所有
  • 客服微信

  • 服务号