Data placement strategy for parallel XML databases

Wang Guo-Ren<sup>*</sup>; Tang Nan; Yu Ya-Xin; Sun Bing; Yu Ge

doi:10.1360/jos170770

摘要

This paper targets on parallel XML document partitioning strategies to process XML queries in parallel. To describe the problem of XML data partitioning, a concept, intermediary node, is presented in this paper. By a set of intermediary nodes, an XML data tree can be partitioned into a root-tree and a set of sub-trees. While the root-tree is duplicated over all the nodes, the set of the sub-trees can be evenly partitioned over all the nodes based on the workload of user queries. For the same XML data tree, there are a number of intermediary nodes sets, and different intermediary nodes sets will generate different partitions. It can be evaluated if a partitioning is good based on the workload of user queries. It is obviously an NP hard problem to choose an optimal partitioning. To solve this problem, this paper proposes a set of heuristic rules. Based on the idea described above, this paper designs and implements an XML data partitioning algorithm, WIN, and the extensive experimental results show that its speedup and scaleup performances outperform the existing strategies.

出版日期2006
单位东北大学

全文

访问全文

收藏分享被引(4) 浏览

更新时间：2022-12-21 18:13

Data placement strategy for parallel XML databases

摘要

全文

产品服务

站内浏览

服务支持

联系方式

科研之友