REX-Reference-Cited by-同舟云学术

REX

Published:2012-07 Issue:11 Volume:5 Page:1280-1291
ISSN:2150-8097
Container-title:Proceedings of the VLDB Endowment
language:en
Short-container-title:Proc. VLDB Endow.

Author:

Mihaylov Svilen R.¹,Ives Zachary G.¹,Guha Sudipto¹

Affiliation:

1. University of Pennsylvania, Philadelphia, PA

Abstract

In today's Web and social network environments, query workloads include ad hoc and OLAP queries, as well as iterative algorithms that analyze data relationships (e.g., link analysis, clustering, learning). Modern DBMSs support ad hoc and OLAP queries, but most are not robust enough to scale to large clusters. Conversely, "cloud" platforms like MapReduce execute chains of batch tasks across clusters in a fault tolerant way, but have too much overhead to support ad hoc queries. Moreover, both classes of platform incur significant overhead in executing iterative data analysis algorithms. Most such iterative algorithms repeatedly refine portions of their answers, until some convergence criterion is reached. However, general cloud platforms typically must reprocess all data in each step. DBMSs that support recursive SQL are more efficient in that they propagate only the changes in each step --- but they still accumulate each iteration's state, even if it is no longer useful. User-defined functions are also typically harder to write for DBMSs than for cloud platforms. We seek to unify the strengths of both styles of platforms, with a focus on supporting iterative computations in which changes , in the form of deltas , are propagated from iteration to iteration, and state is efficiently updated in an extensible way. We present a programming model oriented around deltas, describe how we execute and optimize such programs in our REX runtime system, and validate that our platform also handles failures gracefully. We experimentally validate our techniques, and show speedups over the competing methods ranging from 2.5 to nearly 100 times.

Publisher

VLDB Endowment

Subject

General Earth and Planetary Sciences,Water Science and Technology,Geography, Planning and Development

Link

https://dl.acm.org/doi/pdf/10.14778/2350229.2350246

Cited by 36 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Quantitative estimates of the metachromasia reaction of volutin granules of yeast using neural networks;Artificial Intelligence;2024-06-28

2. Handling Iterations in Distributed Dataflow Systems;ACM Computing Surveys;2022-12-31

3. Toward High-Performance Delta-Based Iterative Processing with a Group-Based Approach;Journal of Computer Science and Technology;2022-07

4. Hybrid Evaluation for Distributed Iterative Matrix Computation;Proceedings of the 2021 International Conference on Management of Data;2021-06-09

5. Formal semantics and high performance in declarative machine learning using Datalog;The VLDB Journal;2021-05-31