发明名称 Very-large-scale automatic categorizer for Web content
摘要 A method and apparatus for efficiently classifying and categorizing data objects such as electronic text, graphics, and audio based documents within very-large-scale hierarchical classification trees is provided. In accordance with one embodiment of the invention, a first node of a plurality of nodes of a subject hierarchy is selected. Previously classified data objects corresponding to a selected first node of a subject hierarchy as well as any associated sub-nodes of the selected node are aggregated to form a content class of data objects. Similarly, data objects corresponding to sibling nodes of the selected node and any associated sub-nodes of the sibling nodes are then aggregated to form an anti-content class of data objects. Features are then extracted from each of the content class of data objects and the anti-content class of data objects to facilitate characterization of said previously classified data objects.
申请公布号 US2005021545(A1) 申请公布日期 2005.01.27
申请号 US20040923431 申请日期 2004.08.20
申请人 MICROSOFT CORPORATION 发明人 LULICH DANIEL P.;GUILAK FARZIN G.
分类号 G06F17/30;(IPC1-7):G06F17/00 主分类号 G06F17/30
代理机构 代理人
主权项
地址
您可能感兴趣的专利