View-based query processing: On the relationship between rewriting, answering and losslessness

被引:24
作者
Calvanese, Diego
De Giacomo, Giuseppe
Lenzerini, Maurizio
Vardi, Moshe Y.
机构
[1] Free Univ Bozen Bolzano, Fac Comp Sci, I-39100 Bolzano, Italy
[2] Univ Roma La Sapienza, Dipartimento Informat & Sistemist, I-00198 Rome, Italy
[3] Rice Univ, Dept Comp Sci, Houston, TX 77251 USA
基金
美国国家科学基金会; 欧盟地平线“2020”;
关键词
query containment; query rewriting; query answering; losslessness; conjunctive queries; regular path queries; semistructured data;
D O I
10.1016/j.tcs.2006.11.006
中图分类号
TP301 [理论、方法];
学科分类号
081202 ;
摘要
As a result of the extensive research in view-based query processing, three notions have been identified as fundamental, namely rewriting, answering, and losslessness. Answering amounts to computing the tuples satisfying the query in all databases consistent with the views. Rewriting consists in first reformulating the query in terms of the views and then evaluating the rewriting over the view extensions. Losslessness holds if we can answer the query by solely relying on the content of the views. While the mutual relationship between these three notions is easy to identify in the case of conjunctive queries, the terrain of notions gets considerably more complicated going beyond such a query class. In this paper, we revisit the notions of answering, rewriting, and losslessness and clarify their relationship in the setting of semistructured databases, and in particular for the basic query class in this setting, i.e., two-way regular path queries. Our first result is a clean explanation of the relationship between answering and rewriting, in which we characterize rewriting as a "linear approximation" of query answering. We show that applying this linear approximation to the constraint-satisfaction framework yields an elegant automata-theoretic approach to query rewriting. As for losslessness, we show that there are indeed two distinct interpretations for this notion, namely with respect to answering, and with respect to rewriting. We also show that the constraint-theoretic approach and the automata-theoretic approach can be combined to give algorithmic characterization of the various facets of losslessness. Finally, we deal with the problem of coping with loss, by considering mechanisms aimed at explaining lossiness to the user. (c) 2006 Elsevier B.V. All rights reserved.
引用
收藏
页码:169 / 182
页数:14
相关论文
共 33 条
  • [21] FERNANDEZ MF, 1998, P ACM SIGMOD INT C M, P414
  • [22] Rewriting queries using views
    Flesca, S
    Greco, S
    [J]. IEEE TRANSACTIONS ON KNOWLEDGE AND DATA ENGINEERING, 2001, 13 (06) : 980 - 995
  • [23] Florescu D., 1998, SIGMOD Record, V27, P59, DOI 10.1145/290593.290605
  • [24] Grahne G, 1999, LECT NOTES COMPUT SC, V1540, P332
  • [25] Answering queries using views: A survey
    Halevy, AY
    [J]. VLDB JOURNAL, 2001, 10 (04) : 270 - 294
  • [26] Hopcroft J. E., 2007, Introduction to Automata Theory, Languages and Computation
  • [27] Levy A. Y., 1995, Proceedings of the Fourteenth ACM SIGACT-SIGMOD-SIGART Symposium on Principles of Database Systems. PODS 1995, P95, DOI 10.1145/212433.220198
  • [28] Li C, 2001, LECT NOTES COMPUT SC, V1973, P99
  • [29] Index structures for path expressions
    Milo, T
    Suciu, D
    [J]. DATABASE THEORY - ICDT'99, 1999, 1540 : 277 - 295
  • [30] Tininini L., 2000, P 19 ACM SIGACT SIGM, P47