An Efficient and Expressive Similarity Measure for Relational Clustering Using Neighbourhood Trees

机译：使用邻域树的关系群集有效和表达的相似度量

获取原文

页面导航

摘要
著录项
相似文献
相关主题

摘要

Clustering is an under-specified task: there are no universal criteria for what makes a good clustering. This is especially true for relational data, where similarity can be based on the features of individuals, the relationships between them, or a mix of both. Existing methods for relational clustering have strong and often implicit biases in this respect. In this paper, we introduce a novel similarity measure for relational data. It is the first measure to incorporate a wide variety of types of similarity, including similarity of attributes, similarity of relational context, and proximity in a hypergraph. We experimentally evaluate how using this similarity affects the quality of clustering on very different types of datasets. The experiments demonstrate that (a) using this similarity in standard clustering methods consistently gives good results, whereas other measures work well only on datasets that match their bias; and (b) on most datasets, the novel similarity outperforms even the best among the existing ones.

机译：群集是一个低于指定的任务：没有通用标准，以使良好的聚类是什么。对于关系数据尤其如此，其中相似性可以基于个人的特征，它们之间的关系或两者的混合。关于关系聚类的现有方法具有很强的且通常隐含的偏见。在本文中，我们介绍了一种用于关系数据的新颖相似度量。它是包含各种类型的相似性的第一次数，包括属性的相似性，关系上下文的相似性，以及超图中的接近度。我们通过实验评估这种相似度如何影响群集的群集在非常不同类型的数据集上。实验表明（a）在标准聚类方法中使用这种相似性始终如一地提供良好的结果，而其他措施仅适用于与其偏见的数据集工作; （b）在大多数数据集上，新颖的相似性甚至是现有的相似性。

著录项

来源
《European Conference on Artificial Intelligence》|2016年|913-1833p|共2页
会议地点
作者
Sebastijan Dumancic; Hendrik Blockeel;
展开▼
作者单位

展开▼
会议组织
原文格式 PDF
正文语种
中图分类 TP18-53;
关键词

相似文献

外文文献
中文文献
专利

1. An expressive dissimilarity measure for relational clustering using neighbourhood trees [J] . Dumancic Sebastijan, Blockeel Hendrik Machine Learning . 2017,第9a10期

机译：使用邻域树的关系聚类的表达差异度量
2. Efficient similarity measure for comparing tree structures [J] . Fatiha Souam, Ali Aiet El Hadj International journal of advanced intelligence paradigms . 2016,第1期

机译：用于比较树结构的有效相似性度量
3. Efficient text document clustering with new similarity measures [J] . R. Lakshmi, S. Baskar International Journal of Business Intelligence and Data Mining . 2021,第1期

机译：具有新的相似度量的高效文本文档聚类
4. An Efficient and Expressive Similarity Measure for Relational Clustering Using Neighbourhood Trees [C] . Sebastijan Dumancic, Hendrik Blockeel European Conference on Artificial Intelligence . 2016

机译：使用邻域树的关系群集有效和表达的相似度量
5. A comparison of clustering procedures and similarity measures in creating clusters using warp functions. [D] . Elguindi, Anne Charlotte. 2010

机译：使用warp函数创建聚类时聚类过程和相似性度量的比较。
6. GO functional similarity clustering depends on similarity measure clustering method and annotation completeness [O] . Meng Liu, Paul D. Thomas 2019

机译：GO功能相似性聚类取决于相似性度量聚类方法和注释完整性
7. An efficient and expressive similarity measure for relational clustering using neighbourhood trees [O] . Dumancic Sebastijan, Blockeel Hendrik 2016

机译：使用邻域树进行关系聚类的有效且表达相似的度量
8. A NEW MEASURE OF BIOTIC SIMILARITY BETWEEN SAMPLES AND ITS APPLICATIONS WITH A CLUSTER ANALYSIS PROGRAM [R] . Carlos F. A. Pinkham 1974

机译：利用聚类分析程序测量样品间的生物相似性及其应用的新方法

An Efficient and Expressive Similarity Measure for Relational Clustering Using Neighbourhood Trees

摘要

著录项

相似文献

相关主题

期刊订阅