面向分布式图计算作业的容错技术研究综述  被引量:4

Survey of State-of-the-art Fault Tolerance for Distributed Graph Processing Jobs

在线阅读下载全文

作  者:张程博 李影[1,2] 贾统 ZHANG Cheng-Bo;LI Ying;JIA Tong(School of Software and Microelectronics,Peking University,Beijing 102600,China;National Engineering Research Center for Software Engineering,Peking University,Beijing 100871,China;School of Electronics Engineering and Computer Science,Peking University,Beijing 100871,China)

机构地区:[1]北京大学软件与微电子学院,北京102600 [2]北京大学软件工程国家工程研究中心,北京100871 [3]北京大学信息科学技术学院,北京100871

出  处:《软件学报》2021年第7期2078-2102,共25页Journal of Software

基  金:广东省重点领域研发计划(2020B010164003)。

摘  要:随着图数据规模的日益庞大和图计算作业的日益复杂,图计算的分布化成为必然趋势.然而图计算作业在运行过程中面临着分布式图计算系统内外各种来源的非确定性所带来的严峻的可靠性问题.首先分析了分布式图计算框架中不确定性因素和不同类型图计算作业的鲁棒性,并提出了基于成本、效率和质量3个维度的面向分布式图计算作业的容错技术评估框架,然后分别对分布式图计算的4种容错机制——基于检查点的容错、基于日志的容错、基于复制的容错、基于算法补偿的容错等机制结合国内外相关工作做了深入的分析、评估和比较.最后对未来的研究方向进行了展望.As the growth of graph data scale and complexity of graph processing,the trend of distributed graph processing shall be inevitable.However,graph processing jobs run with severe reliability problems caused by the uncertainty originated from inside and outside the distributed graph processing system.This study first analyzes the uncertainty factors of the distributed graph processing frameworks and the robustness of different types of graph processing jobs;then proposes an evaluation framework of fault tolerance for distributed graph processing based on cost,efficiency,and quality of fault tolerance.This study also analyzes,evaluates,and compares the four fault-tolerant mechanisms of distributed graph processing-checkpointing based fault tolerance,logging based fault tolerance,replication based fault tolerance,and algorithm compensation based fault tolerance-combining related researches.Finally,the direction of future researches is prospected.

关 键 词:图数据 故障和失效 分布式图计算 容错机制 非确定性软件系统 

分 类 号:TP311[自动化与计算机技术—计算机软件与理论]

 

参考文献:

正在载入数据...

 

二级参考文献:

正在载入数据...

 

耦合文献:

正在载入数据...

 

引证文献:

正在载入数据...

 

二级引证文献:

正在载入数据...

 

同被引文献:

正在载入数据...

 

相关期刊文献:

正在载入数据...

相关的主题
相关的作者对象
相关的机构对象