首页>
外国专利>
VERY-LARGE-SCALE AUTOMATIC CATEGORIZER FOR WEB CONTENT
VERY-LARGE-SCALE AUTOMATIC CATEGORIZER FOR WEB CONTENT
展开▼
机译:用于Web内容的超大型自动分类器
展开▼
页面导航
摘要
著录项
相似文献
摘要
A method and apparatus for efficiently classifying and categorizing data objects such as electronic text, graphics, and audio based documents within very-large-scale hierarchical classification trees is provided. In accordance with one embodiment of the invention, a first node of a plurality of nodes of a subject hierarchy is selected. Previously classified data objects (202) corresponding to a selected first node of a subject hierarchy as well as any associated sub-nodes of the selected node are aggregated to form a content class of data objects(206). Similarly, data objects corresponding to sibling nodes of the selected node and any associated sub-nodes of the sibling nodes are then aggregated to form an anti-content class of data objects (206). Features are then extracted (207) from each of the content class of data objects and the anti-content class of data objects to facilitate characterization (209) of said previously classified data objects (202).
展开▼