检索规则说明:AND代表“并且”;OR代表“或者”;NOT代表“不包含”;(注意必须大写,运算符两边需空一格)
检 索 范 例 :范例一: (K=图书馆学 OR K=情报学) AND A=范并思 范例二:J=计算机应用与软件 AND (U=C++ OR U=Basic) NOT M=Visual
作 者:张福新[1] 章隆兵[1] 胡伟武[1] 唐志敏[1]
出 处:《软件学报》2005年第2期165-173,共9页Journal of Software
基 金:国家自然科学基金~~
摘 要:软件 DSM(distributed shared memory)系统在机群上构造了共享存储编程环境,结合了共享存储的易编程性和机群的可扩展性,引起了广泛的研究.由于软件 DSM 系统是一个分布式系统,系统失败风险大,需要实现容错技术以促进其实用化.利用用户级检查点技术,在支持域存储一致模型的软件 DSM 系统 JIAJIA 的基础上,设计并实现了一个可恢复的高可移植的软件 DSM 系统 JIACKPT(JIAjia with ChecKPoinTing).由于采用适合软件 DSM 系统的强全局一致状态以及多种优化措施,JIACKPT 易于实现且获得很好的性能.在一个 8 节点的 PC 机群上的应用测试表明,即使每分钟做一次检查点,大部分应用的检查点开销也小于 10%.此外,JIACKPT 还具有高可移植性.这些都表明 JIACKPT 已经成为一个比较实用的系统.Software distributed shared memory (DSM) system has constructed a virtual shared memory abstract on cluster, which combines the programmability of shared memory and fine scalability of cluster. So it is widely studied. Software DSM system is easy to fail because it is a distributed system, some kinds of fault tolerance are necessary for it to be more practical. A recoverable and portable software DSM system, JIACKPT (JIAjia with ChecKPoinTing), has been designed and implemented to tolerate the fault of system. JIACKPT, based on JIAJIA, has adopted the checkpointing technology. By maintaining the strict global consistent state and using some optimization techniques, JIACKPT has gotten high performance. The experimental results on an 8-node PC cluster show that the checkpoint overhead is less than 10% of the whole execution time when checkpoint is done once per minute. JIACKPT also has good portability and can run on several operating systems, such as Linux, Solaris, etc. JIACKPT is a practical recoverable software DSM system.
关 键 词:软件DSM系统 检查点 全局一致状态 JIAJIA
分 类 号:TP311[自动化与计算机技术—计算机软件与理论]
正在载入数据...
正在载入数据...
正在载入数据...
正在载入数据...
正在载入数据...
正在载入数据...
正在载入数据...
正在链接到云南高校图书馆文献保障联盟下载...
云南高校图书馆联盟文献共享服务平台 版权所有©
您的IP:216.73.216.33