首页> 外文会议>IEEE/WIC/ACM Joint International Conference on Web Intelligence and Intelligent Agent Technology >Nowhere to Hide: Finding Plagiarized Documents Based on Sentence Similarity
【24h】

Nowhere to Hide: Finding Plagiarized Documents Based on Sentence Similarity

机译:无处可隐藏:根据句子相似寻找抄袭文档

获取原文

摘要

Plagiarism is a serious problem that infringes copyrighted documents/materials, which is an unethical practice and decreases the economic incentive received by authors (owners) of the original copies. Unfortunately, plagiarism is getting worse due to the increasing number of on-line publications on the Web, which facilitates locating and paraphrasing information. In solving this problem, we propose a novel plagiarism-detection method, called SimPaD, which (i) establishes the degree of resemblance between any two documents D1 and D2 based on their sentence-to-sentence similarity computed by using pre-defined word-correlation factors, and (ii) generates agraphical view of sentences that are similar (or the same) in D1 and D2. Experimental results verify that SimPaD is highly accurate in detecting (non-) plagiarized documents and outperforms existing plagiarism-detection approaches.
机译:抄袭是侵犯受版权保护的文件/材料的严重问题,这是一个不道德的实践,并降低了原始副本的作者(业主)所收回的经济激励。不幸的是,由于网络上越来越多的在线出版物,抄袭是越来越糟的,这有助于定位和解释信息。在解决这个问题时,我们提出了一种名为SIMPAD的新型抄袭检测方法,(i)基于通过使用预定定义的单词计算的句子相似度,建立了任何两个文档D1和D2之间的相似度程度相关因子和(ii)生成句子的Agraphical视图,在D1和D2中类似(或相同)。实验结果验证SIMPAD在检测(非)抄袭文件中高度准确,优于现有的抄袭检测方法。

著录项

相似文献

  • 外文文献
  • 中文文献
  • 专利
获取原文

客服邮箱:kefu@zhangqiaokeyan.com

京公网安备:11010802029741号 ICP备案号:京ICP备15016152号-6 六维联合信息科技 (北京) 有限公司©版权所有
  • 客服微信

  • 服务号