Sector and Sphere: the design and implementation of a high-performance data cloud

被引:62
作者
Gu, Yunhong
Grossman, Robert L. [1 ,2 ]
机构
[1] Univ Illinois, Natl Ctr Data Min, Chicago, IL 60607 USA
[2] Open Data Grp, River Forest, IL 60305 USA
来源
PHILOSOPHICAL TRANSACTIONS OF THE ROYAL SOCIETY A-MATHEMATICAL PHYSICAL AND ENGINEERING SCIENCES | 2009年 / 367卷 / 1897期
基金
美国国家科学基金会;
关键词
cloud computing; data-intensive computing; distributed computing;
D O I
10.1098/rsta.2009.0053
中图分类号
O [数理科学和化学]; P [天文学、地球科学]; Q [生物科学]; N [自然科学总论];
学科分类号
07 ; 0710 ; 09 ;
摘要
Cloud computing has demonstrated that processing very large datasets over commodity clusters can be done simply, given the right programming model and infrastructure. In this paper, we describe the design and implementation of the Sector storage cloud and the Sphere compute cloud. By contrast with the existing storage and compute clouds, Sector can manage data not only within a data centre, but also across geographically distributed data centres. Similarly, the Sphere compute cloud supports user-defined functions (UDFs) over data both within and across data centres. As a special case, MapReduce-style programming can be implemented in Sphere by using a Map UDF followed by a Reduce UDF. We describe some experimental studies comparing Sector/Sphere and Hadoop using the Terasort benchmark. In these studies, Sector is approximately twice as fast as Hadoop. Sector/Sphere is open source.
引用
收藏
页码:2429 / 2445
页数:17
相关论文
共 19 条
[1]  
[Anonymous], 19 ACM S OP SYST PRI
[2]  
Babcock B., 2002, PODS, P1, DOI [DOI 10.1145/543613.543615, 10.1145/543613.543615]
[3]  
BEYNON MD, 2000, MASS STOR SYST C COL
[4]  
Borthaku D., 2007, The Hadoop distributed file system
[5]  
Chang F., 2006, OSDI 06
[6]  
CHEN L., 2004, 13 IEEE INT S HIGH P
[7]  
Dean J, 2004, OSDI, P137
[8]   Globally distribued content delivery [J].
Dilley, J ;
Maggs, B ;
Parikh, J ;
Prokop, H ;
Sitaraman, R ;
Weihl, B .
IEEE INTERNET COMPUTING, 2002, 6 (05) :50-58
[9]  
Foster I, 2005, LECT NOTES COMPUT SC, V3779, P2
[10]  
Gedik Bugra., 2007, VLDB 07, P1286