Ιδρυματικό Αποθετήριο
Πολυτεχνείο Κρήτης
EN  |  EL

Αναζήτηση

Πλοήγηση

Ο Χώρος μου

Sketch-based querying of distributed sliding-window data streams

Papapetrou Odysseas, Garofalakis Minos, Deligiannakis Antonios

Απλή Εγγραφή


URIhttp://purl.tuc.gr/dl/dias/73B079AE-B219-4669-99D7-C156B1AFB8C3-
Αναγνωριστικόhttp://vldb.org/pvldb/vol5/p992_odysseaspapapetrou_vldb2012.pdf-
Αναγνωριστικόhttps://doi.org/10.14778/2336664.2336672-
Γλώσσαen-
Μέγεθος12 pagesen
ΤίτλοςSketch-based querying of distributed sliding-window data streamsen
ΔημιουργόςPapapetrou Odysseasen
ΔημιουργόςΠαπαπετρου Οδυσσεαςel
ΔημιουργόςGarofalakis Minosen
ΔημιουργόςΓαροφαλακης Μινωςel
ΔημιουργόςDeligiannakis Antoniosen
ΔημιουργόςΔεληγιαννακης Αντωνιοςel
ΕκδότηςAssociation for Computing Machineryen
ΠερίληψηWhile traditional data-management systems focus on evaluating single, adhoc queries over static data sets in a centralized setting, several emerging applications require (possibly, continuous) answers to queries on dynamic data that is widely distributed and constantly updated. Furthermore, such query answers often need to discount data that is “stale”, and operate solely on a sliding window of recent data arrivals (e.g., data updates occurring over the last 24 hours). Such distributed data streaming applications mandate novel algorithmic solutions that are both time- and space-efficient (to manage high-speed data streams), and also communication-efficient (to deal with physical data distribution). In this paper, we consider the problem of complex query answering over distributed, high-dimensional data streams in the sliding-window model. We introduce a novel sketching technique (termed ECM-sketch) that allows effective summarization of streaming data over both time-based and count-based sliding windows with probabilistic accuracy guarantees. Our sketch structure enables point as well as inner-product queries, and can be employed to address a broad range of problems, such as maintaining frequency statistics, finding heavy hitters, and computing quantiles in the sliding-window model. Focusing on distributed environments, we demonstrate how ECM-sketches of individual, local streams can be composed to generate a (low-error) ECM-sketch summary of the order-preserving aggregation of all streams; furthermore, we show how ECM-sketches can be exploited for continuous monitoring of sliding-window queries over distributed streams. Our extensive experimental study with two real-life data sets validates our theoretical claims and verifies the effectiveness of our techniques. To the best of our knowledge, ours is the first work to address efficient, guaranteed-error complex query answering over distributed data streams in the sliding-window model. en
ΤύποςΠλήρης Δημοσίευση σε Συνέδριοel
ΤύποςConference Full Paperen
Άδεια Χρήσηςhttp://creativecommons.org/licenses/by/4.0/en
Ημερομηνία2015-11-30-
Ημερομηνία Δημοσίευσης2012-
Θεματική ΚατηγορίαInformation systemsen
Θεματική ΚατηγορίαData managementen
Βιβλιογραφική ΑναφοράO. Papapetrou, M. Garofalakis and A. Deligiannakis, "Sketch-based querying of distributed sliding-window data streams", in 2012 VLDB Endowment, vol. 5, no. 10, pp. 992-1003. doi: 10.14778/2336664.2336672 en

Υπηρεσίες

Στατιστικά