Deakin University
Browse

File(s) under permanent embargo

A highly practical approach toward achieving minimum data sets storage cost in the cloud

journal contribution
posted on 2013-06-01, 00:00 authored by D Yuan, Y Yang, Xiao LiuXiao Liu, W Li, L Cui, M Xu, J Chen
Massive computation power and storage capacity of cloud computing systems allow scientists to deploy computation and data intensive applications without infrastructure investment, where large application data sets can be stored in the cloud. Based on the pay-as-you-go model, storage strategies and benchmarking approaches have been developed for cost-effectively storing large volume of generated application data sets in the cloud. However, they are either insufficiently cost-effective for the storage or impractical to be used at runtime. In this paper, toward achieving the minimum cost benchmark, we propose a novel highly cost-effective and practical storage strategy that can automatically decide whether a generated data set should be stored or not at runtime in the cloud. The main focus of this strategy is the local-optimization for the tradeoff between computation and storage, while secondarily also taking users' (optional) preferences on storage into consideration. Both theoretical analysis and simulations conducted on general (random) data sets as well as specific real world applications with Amazon's cost model show that the cost-effectiveness of our strategy is close to or even the same as the minimum cost benchmark, and the efficiency is very high for practical runtime utilization in the cloud.

History

Journal

IEEE transactions on parallel and distributed systems

Volume

24

Issue

6

Pagination

1234 - 1244

Publisher

Institute of Electrical and Electronics Engineers

Location

Piscataway, N.J.

ISSN

1045-9219

Language

eng

Publication classification

C1.1 Refereed article in a scholarly journal; C Journal article

Copyright notice

2013, IEEE