Probabilistic wavelet synopses-Reference-Cited by-同舟云学术

Probabilistic wavelet synopses

Published:2004-03 Issue:1 Volume:29 Page:43-90
ISSN:0362-5915
Container-title:ACM Transactions on Database Systems
language:en
Short-container-title:ACM Trans. Database Syst.

Author:

Garofalakis Minos¹,Gibbons Phillip B.²

Affiliation:

1. Bell Labs, Lucent Technologies, Murray Hill, New Jersey

2. Intel Research, Pittsburgh, Pennsylvania, PA

Abstract

Recent work has demonstrated the effectiveness of the wavelet decomposition in reducing large amounts of data to compact sets of wavelet coefficients (termed "wavelet synopses") that can be used to provide fast and reasonably accurate approximate query answers. A major shortcoming of these existing wavelet techniques is that the quality of the approximate answers they provide varies widely, even for identical queries on nearly identical values in distinct parts of the data. As a result, users have no way of knowing whether a particular approximate answer is highly-accurate or off by many orders of magnitude. In this article, we introduce Probabilistic Wavelet Synopses , the first wavelet-based data reduction technique optimized for guaranteed accuracy of individual approximate answers. Whereas previous approaches rely on deterministic thresholding for selecting the wavelet coefficients to include in the synopsis, our technique is based on a novel, probabilistic thresholding scheme that assigns each coefficient a probability of being included based on its importance to the reconstruction of individual data values, and then flips coins to select the synopsis. We show how our scheme avoids the above pitfalls of deterministic thresholding, providing unbiased , highly accurate answers for individual data values in a data vector. We propose several novel optimization algorithms for tuning our probabilistic thresholding scheme to minimize desired error metrics. Experimental results on real-world and synthetic data sets evaluate these algorithms, and demonstrate the effectiveness of our probabilistic wavelet synopses in providing fast, highly accurate answers with improved quality guarantees.

Publisher

Association for Computing Machinery (ACM)

Subject

Information Systems

Link

https://dl.acm.org/doi/pdf/10.1145/974750.974753

Reference22 articles.

1. Join synopses for approximate query answering

2. Improving responsiveness for wide-area data access;Amsaleg L.;IEEE Data Eng. Bull.,1997

3. Monotone and probabilistic wavelet approximation

4. Probabilistic discrete wavelet approximation

Cited by 47 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. An efficient multidimensional $L_{\infty }$ wavelet method and its application to approximate query processing;World Wide Web;2020-10-10

2. An online PLA algorithm with maximum error bound for generating optimal mixed-segments;International Journal of Machine Learning and Cybernetics;2019-12-20

3. Two-dimensional wavelet synopses with maximum error bound and its application in parallel compression;Journal of Intelligent & Fuzzy Systems;2019-10-09

4. Efficient two-dimensional Haar$$^+$$ synopsis construction for the maximum absolute error measure;The VLDB Journal;2019-07-16

5. Image Scaling: How Hard Can it Be?;IEEE Access;2019