首页> 中文期刊> 《计算机工程与应用》 >突发事件热点话题识别系统及关键问题研究

突发事件热点话题识别系统及关键问题研究

         

摘要

Concerning the system of hot topics detection about the emergency events,an overall technical framework is established to implement the system.Description and solution strategy about the key issues in the four components of the system are provided. In terms of the content and structure features of the news reports as well as the distribution feature of the report sources,the text clipping method and the modified model of feature weighting calculation are proposed based on the VSM text representation model and the TF-IDF formula.The news reports about the earthquake emergency event are evaluated for this model as the data sources.Experimental results indicate that the information such as the headline,the lead and relevant feature parameters by clipping the main body of the news report can be considered as the sample set of the hot topics to be identifiedFurthermore,compared with the classical model,the modified feature items weighting calculation model is more efficient in execution and more adaptive in terms of the text representation capability.%针对突发事件热点话题识别系统,建立了系统实现的整体技术框架,给出了系统四个组成部分的关键问题描述及解决策略,结合新闻报道文本内容和结构的特点和报道源分布性特征,基于VSM文本表示模型和TF-IDF公式,提出了正文裁剪方法和特征权重计算的改进模型,并以地震突发事件新闻报道作为数据源进行模型评估.实验结果表明通过对新闻报道正文的裁剪,只提取标题、导语及相关特征参量等信息即可作为热点话题识别的样本集,且改进的特征权重计算模型与经典模型比较,具有更好地执行效率和适应性更强的文本表示能力.

著录项

相似文献

  • 中文文献
  • 外文文献
  • 专利
获取原文

客服邮箱:kefu@zhangqiaokeyan.com

京公网安备:11010802029741号 ICP备案号:京ICP备15016152号-6 六维联合信息科技 (北京) 有限公司©版权所有
  • 客服微信

  • 服务号