RMIT University
Browse

Efficient evaluation of generalized tree-pattern queries with same-path constraints

conference contribution
posted on 2024-10-31, 16:16 authored by Xiaoying Wu, Dimitri Theodoratos, S Souldatos, T Dalamagas, Timoleon Sellis
Querying XML data is based on the specification of structural patterns which in practice are formulated using XPath. Usually, these structural patterns are in the form of trees (Tree-Pattern Queries - TPQs). Requirements for flexible querying of XML data including XML data from scientific applications have motivated recently the introduction of query languages that are more general and flexible than TPQs. These query languages correspond to a fragment of XPath larger than TPQs for which efficient non-main-memory evaluation algorithms are not known. In this paper, we consider a query language, called Partial Tree-Pattern Query (PTPQ) language, which generalizes and strictly contains TPQs. PTPQs represent a broad fragment of XPath which is very useful in practice. We show how PTPQs can be represented as directed acyclic graphs augmented with "same-path" constraints. We develop an original polynomial time holistic algorithm for PTPQs under the inverted list evaluation model. To the best of our knowledge, this is the first algorithm to support the evaluation of such a broad structural fragment of XPath. We provide a theoretical analysis of our algorithm and identify cases where it is asymptotically optimal. In order to assess its performance, we design two other techniques that evaluate PTPQs by exploiting the state-of-the-art existing algorithms for smaller classes of queries. An extensive experimental evaluation shows that our holistic algorithm outperforms the other ones.

History

Related Materials

  1. 1.
    DOI - Is published in 10.1007/978-3-642-02279-1_27
  2. 2.
    ISSN - Is published in 03029743

Start page

361

End page

379

Total pages

19

Outlet

Proceedings of the 21st International Conference on Scientific and Statistical Database Management (SSDBM 2009)

Editors

Marianne Winslett

Name of conference

21st International Conference on Scientific and Statistical Database Management (SSDBM 2009)

Publisher

Springer

Place published

Berlin, Germany

Start date

2009-06-02

End date

2009-06-04

Language

English

Copyright

© 2009 Springer Berlin Heidelberg

Former Identifier

2006036076

Esploro creation date

2020-06-22

Fedora creation date

2013-01-21