Abstract
Experimental sciences create vast amounts of data. In astronomy, data produced by the Pan-STARRS project (Pan-STARRS project, 2010; Jedicke, Magnier, Kaiser, & Chambers, 2006) is expected to result in more than a petabyte of images every year. In high-energy physics, the Large Hadron Collider will generate 50100 petabytes of data each year, with about 20PB of that data being stored and processed on a worldwide federation of national grids linking 100,000 CPUs (Large Hadron Collider project, 2010; Massimo Lammana, 2004). Cloud computing is immensely appealing to the scientific community, who increasingly see it as being part of the solution to cope with burgeoning data volumes. Cloud computing enables economies-of-scale in facility design and hardware construction. Groups of users are allowed to host, process, and analyze large volumes of data from various sources. There are several vendors that offer cloud computing platforms; these include Amazon Web Services (2010), Google's App Engine (2010), AT&T's Synaptic Hosting (2010), Rackspace (2010), GoGrid (2010) and AppNexus (2010). These vendors promise seemingly infinite amounts of computing power and storage that can be made available on demand, in a pay-only-for-what-you-use pricing model.
Cite
CITATION STYLE
Pallickara, S. L., Pallickara, S., & Pierce, M. (2010). Scientific Data Management in the Cloud: A Survey of Technologies, Approaches and Challenges. In Handbook of Cloud Computing (pp. 517–533). Springer US. https://doi.org/10.1007/978-1-4419-6524-0_22
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.