发明申请
US20100306273A1 APPARATUS, SYSTEM, AND METHOD FOR EFFICIENT CONTENT INDEXING OF STREAMING XML DOCUMENT CONTENT
失效
用于流媒体XML文档内容的高效内容索引的设备,系统和方法
- 专利标题: APPARATUS, SYSTEM, AND METHOD FOR EFFICIENT CONTENT INDEXING OF STREAMING XML DOCUMENT CONTENT
- 专利标题(中): 用于流媒体XML文档内容的高效内容索引的设备,系统和方法
-
申请号: US12475999申请日: 2009-06-01
-
公开(公告)号: US20100306273A1公开(公告)日: 2010-12-02
- 发明人: James P. Branigan , David P. Charboneau , Simon K. Johnston
- 申请人: James P. Branigan , David P. Charboneau , Simon K. Johnston
- 申请人地址: US NY Armonk
- 专利权人: International Business Machines Corporation
- 当前专利权人: International Business Machines Corporation
- 当前专利权人地址: US NY Armonk
- 主分类号: G06F17/30
- IPC分类号: G06F17/30
摘要:
An apparatus, system, and method are disclosed for efficient content indexing of streaming XML document content. A forest generator generates an XML pattern forest from a set of structured index path expressions, the XML pattern forest includes trees and twigs generated from structured index path expressions uniquely associated with a namespace indicator for an XML node. The XML node is identified in a stream of at least one XML document. A comparison module compares the XML node to nodes of trees and twigs of the XML pattern forest. A determination module determines a match between the XML node and an index node in one of a tree and a twig of the XML pattern forest. The index node has a path from an ancestor node to the index node that matches the axis steps of at least one of the structured index path expressions. A storage module stores an index entry for the XML node in response to the determined match, the index entry includes a XML document identifier, an XML node name, a namespace indicator for the XML node, and XML node content.
公开/授权文献
信息查询