基于维基百科的领域概念语义知识库的自动构建方法

张巧燕; 林民; 张树钧

首页> 中文期刊> 《计算机应用研究》 >基于维基百科的领域概念语义知识库的自动构建方法

基于维基百科的领域概念语义知识库的自动构建方法

AI论文写作 >>

开具论文收录证明 >>

页面导航

摘要
著录项
引文网络
相似文献
相关主题

摘要

针对为检索服务的语义知识库存在的内容不全面和不准确的问题,提出一种基于维基百科的软件工程领域概念语义知识库的构建方法.以SWEBOK V3概念为标准,从维基百科提取概念的解释文本,并抽取其关键词表示概念的语义;通过概念在维基百科中的层次关系、概念与其他概念的解释文本关键词之间的链接关系、不同概念解释文本关键词之间的链接关系构成概念语义知识库;利用LDA主题模型分别与TF-IDF、TextRank算法相结合的两种方法抽取关键词;对构建好的概念语义知识库用随机游走算法计算概念间的语义相似度.将实验结果与人工标注结果对比后发现,本方法构建的语义知识库语义相似度准确率能够达到84％以上,充分验证了所提方法的有效性.%The problem of incomplete and inaccurate content for the retrieval of semantic knowledge base existed,this paper proposed a method of constructing the concept semantic knowledge base in the field of software engineering based on Wikipedia.First,taking the concept of SWEBOK V3 as the standard,it extracted the interpretation of the concept from Wikipedia and extracted the keywords to represent the semantic meaning of the concept.Second,through hierarchical relationships of the concept in Wikipedia,link relationships between concepts and explanatory text of other concepts and link relationships between explanatory texts of different concepts,it built concept semantic knowledge base.Then,it combined the LDA topic model with the two methods that were called TF-IDF algorithm and TextRank algorithm respectively serve the keywords extraction.Finally,it calculated the semantic similarity between concepts by the random walk algorithm for the construction of the concept semantic knowledge base.The experimental results were compared with the manual annotation results.The semantic similarity of knowledge base constructed by this method can reach more than 84％.The effectiveness of the proposed method is verified.

著录项

来源
《计算机应用研究》 |2018年第1期|130-134,139|共6页
作者
张巧燕; 林民; 张树钧;
展开▼
作者单位

内蒙古师范大学计算机与信息工程学院;

呼和浩特010022;

内蒙古师范大学计算机与信息工程学院;

呼和浩特010022;

内蒙古师范大学计算机与信息工程学院;

呼和浩特010022;

展开▼
原文格式 PDF
正文语种 chi
中图分类文字信息处理;
关键词
维基百科; 语义知识库; 关键词抽取; 语义相似度计算; 随机游走;

相似文献

中文文献
外文文献
专利

1. 基于维基百科的领域本体自动构建方法研究 [J] . 吴洁明 ,刘雁昆 ,段建勇 . 计算机应用与软件 . 2016,第007期
2. 基于维基百科的语义知识库及其构建方法研究 [J] . 张海粟 ,马大明 ,邓智龙 . 计算机应用研究 . 2011,第008期
3. 基于领域语义知识库的疾病辅助诊断方法 [J] . 陈德彦 ,赵宏 ,张霞 . 软件学报 . 2020,第010期
4. 基于维基百科网络技术的概念语义网络构建 [J] . 杨建萍 ,年梅 . 计算机与现代化 . 2016,第001期
5. 基于维基百科的领域概念知识点自动问答系统的设计与实现 [J] . 张巧燕 ,裴栋 ,薛慧君 . 电脑编程技巧与维护 . 2021,第004期
6. 一种基于语义关系计算领域本体中概念间语义相关度的方法 [C] . 田萱 ,杜小勇 ,李海华 . 第二十四届中国数据库学术会议 . 2007
7. 基于维基百科构建语义知识库及其在文本分类领域的应用研究 [A] . 苏小康 . 2010

基于维基百科的领域概念语义知识库的自动构建方法

摘要

著录项

引文网络

相似文献

相关主题

期刊订阅