首页>
外国专利>
Device, method and program for generating accurate corpus data for presentation target for searching
Device, method and program for generating accurate corpus data for presentation target for searching
展开▼
机译:用于生成用于呈现目标以供搜索的准确语料数据的装置,方法和程序
展开▼
页面导航
摘要
著录项
相似文献
摘要
A corpus generation device according to an embodiment includes a web page acquisition unit, a reference word acquisition unit, an attachment unit and an output unit. The web page acquisition unit acquires a web page including description sentence data regarding a presentation target. The reference word acquisition unit acquires a reference word that is an attribute value regarding the presentation target from the web page. The attachment unit extracts a broader word belonging to a layer above the reference word acquired by the reference word acquisition unit from a storage unit that stores hierarchical relationship information indicating a hierarchical relationship between attribute values, and attaches an attribute tag corresponding to the reference word to the broader word included in the description sentence data. The output unit outputs, as corpus data, the description sentence data to which the attribute tag is attached by the attachment unit.
展开▼