An Optimal Checkpointing Model with Online OCI Adjustment for Stream Processing Applications.

CONCURRENCY AND COMPUTATION-PRACTICE & EXPERIENCE(2019)

引用 15|浏览41
暂无评分
摘要
Checkpoint-based fault-tolerant (FT) methods have been widely used to enhance the reliability of stream processing systems, but a checkpointing process usually introduces considerable overhead. It is a critical issue to choose the optimal checkpoint interval (OCI) that maximizes the processing efficiency. Traditional OCI models consider the recovery time equals to the execution time from the last checkpoint to the failure moment. However, for stream processing jobs, the recovery time is related to reprocessing workloads, depending on the real-time input data before a failure. A new model is needed to choose the OCI for stream processing applications. Moreover, the input data rate of a stream processing job fluctuates over time. To solve these problems, we present a novel DSPS OCI (DOCI) model in this paper. We prove that it maximizes the processing efficiency for a given time. We propose an approach to dynamically adjust the OCI for an application to accommodate the workload fluctuations. We conduct simulation experiments to verify the effectiveness of our DOCI model and the efficiency of the online OCI adjustment algorithm. Experimental results with a real-world dataset show that DOCI achieves an improvement on system efficiency by up to 32%, compared with existing FT approaches.
更多
查看译文
关键词
distributed stream processing,fault tolerance,optimal checkpoint interval
AI 理解论文
溯源树
样例
生成溯源树,研究论文发展脉络
Chat Paper
正在生成论文摘要