Dynamic data replication and placement strategy in geographically distributed data centers

被引:4
|
作者
Bouhouch, Laila [1 ]
Zbakh, Mostapha [1 ]
Tadonki, Claude [2 ]
机构
[1] Mohammed V Univ Rabat, Natl Sch Comp Sci & Syst Anal, Rabat, Morocco
[2] MINES ParisTech PSL CRI, Paris, France
来源
关键词
big data; cloud computing; Cloudsim; data placement; dynamic data replication; CLOUD; OPPORTUNITIES;
D O I
10.1002/cpe.6858
中图分类号
TP31 [计算机软件];
学科分类号
081202 ; 0835 ;
摘要
With the evolution of geographically distributed data centers in the Cloud Computing landscape along with the amount of data being processed in these data centers, which is growing at an exponential rate, processing massive data applications become an important topic. Since a given task may require many datasets for its execution and the datasets are spread over several different data centers, finding an efficient way to manage the datasets storage across nodes of a Cloud system is a difficult problem. In fact, the execution time of a task might be influenced by the cost of data transfers, which mainly depends on two criterias. The first one is the initial placement of the input datasets during the build-time phase, while the second is the replication of the datasets during the runtime phase. The replication is explicitly considered when datasets are being migrated over the data centers in order to make them locally available wherever needed. Data placement and data replication are important challenges in Cloud Computing. Nevertheless, many studies focus on data placement or data replication exclusively. In this paper, a combination of a data placement strategy followed by a dynamic data replication management strategy is proposed, with the purpose of reducing the associated cost of all data transfers between the (distant) data centers. Our proposed data placement approach considers the main characteristics of a data center such as storage capacity and read/write speeds to efficiently store the datasets, while our dynamic data replication management approach considers three parameters: the number of replicas in the system, the dependency between datasets and tasks and the storage capacity of data centers. The decision of when and whether to keep or to delete replicas is determined by the fulfillment of those three parameters. Our approach estimates the total execution time of the tasks as well as the monetary cost, considering the data transfers activity. Our experiments are conducted using Cloudsim simulator. The obtained results show that our proposed strategies produce an efficient data management by reducing the overheads of the data transfers, compared to both a data placement without replication (by 76%) and the selected data replication approach from Kouidri et al. (by 52%), and by improving the financial cost.
引用
收藏
页数:20
相关论文
共 50 条
  • [31] A study of dynamic data placement for ATLAS distributed data management
    Beermann, T.
    Stewart, G. A.
    Maettig, P.
    21ST INTERNATIONAL CONFERENCE ON COMPUTING IN HIGH ENERGY AND NUCLEAR PHYSICS (CHEP2015), PARTS 1-9, 2015, 664
  • [32] Cost Optimization for Dynamic Replication and Migration of Data in Cloud Data Centers
    Mansouri, Yaser
    Toosi, Adel Nadjaran
    Buyya, Rajkumar
    IEEE TRANSACTIONS ON CLOUD COMPUTING, 2019, 7 (03) : 705 - 718
  • [33] An Optimal Task Placement Strategy in Geo-Distributed Data Centers Involving Renewable Energy
    Wang, Ran
    Lu, Yiwen
    Zhu, Kun
    Hao, Jie
    Wang, Ping
    Cao, Yue
    IEEE ACCESS, 2018, 6 : 61948 - 61958
  • [34] A Data Placement Strategy for Distributed Document-oriented Data Warehouse
    Khalil, Abdelhak
    Belaissaoui, Mustapha
    Toufik, Fouad
    IAENG International Journal of Computer Science, 2023, 50 (04)
  • [35] A dynamic data grid replication strategy to minimize the data missed
    Lei, Ming
    Vrbsky, Susan V.
    Hong, Xiaoyan
    2006 3RD INTERNATIONAL CONFERENCE ON BROADBAND COMMUNICATIONS, NETWORKS AND SYSTEMS, VOLS 1-3, 2006, : 721 - +
  • [36] A Heat-Recirculation-Aware Data Placement Strategy towards Data Centers
    Zhong, Zijie
    Deng, Yuhui
    Li, Jie
    2022 IEEE 28TH INTERNATIONAL CONFERENCE ON PARALLEL AND DISTRIBUTED SYSTEMS, ICPADS, 2022, : 578 - 585
  • [37] Dynamic Placement of Virtualized Resources for Data Centers in Cloud
    Usmin, S.
    Irudayaraja, M. Arockia
    Muthaiah, U.
    2014 INTERNATIONAL CONFERENCE ON INFORMATION COMMUNICATION AND EMBEDDED SYSTEMS (ICICES), 2014,
  • [38] Dynamic strategy of placement of the replicas in data grid
    Belalem, Ghalem
    Bouhraoua, Farouk
    PARALLEL COMPUTING TECHNOLOGIES, PROCEEDINGS, 2007, 4671 : 496 - +
  • [39] Dynamic Data Replication Strategy in Cloud Environments
    Jayalakshmi, D. S.
    Ranjana, Rashmi T. P.
    Srinivasan, R.
    2015 FIFTH INTERNATIONAL CONFERENCE ON ADVANCES IN COMPUTING AND COMMUNICATIONS (ICACC), 2015, : 102 - 105
  • [40] New dynamic replication strategy for data grid
    College of Computer and Communication Engineering, China University of Petroleum , Dongying 257061, China
    不详
    不详
    Beijing Jiaotong Daxue Xuebao, 2008, 6 (111-115+122):