Understanding Data Characteristics and Access Patterns in a Cloud Storage System

被引:25
|
作者
Liu, Songbin [1 ]
Huang, Xiaomeng [1 ]
Fu, Haohuan [1 ]
Yang, Guangwen [1 ]
机构
[1] Tsinghua Univ, Minist Educ, Key Lab Earth Syst Modeling, Beijing 100084, Peoples R China
关键词
Cloud Storage; File System; Data Characteristic; Access Pattern;
D O I
10.1109/CCGrid.2013.11
中图分类号
TP301 [理论、方法];
学科分类号
081202 ;
摘要
Understanding the inherent system characteristics is crucial to the design and optimization of cloud storage system, and few studies have systematically investigated its data characteristics and access patterns. This paper presents an analysis of file system snapshot and five-month access trace of a campus cloud storage system that has been deployed on Tsinghua campus for three years. The system provides online storage and data sharing services for more than 19,000 students and 500 student groups. We report several data characteristics including file size and file type, as well as some access patterns, including read/write ratio, read-write dependency and daily traffic. We find that there are many differences between cloud storage system and traditional file systems: our cloud storage system has larger file sizes, lower read/ write ratio, and smaller set of active files than those of a typical traditional file system. With a trace-driven simulation, we find that the cache efficiency can be improved by 5 times using the guidance from our observations.
引用
收藏
页码:327 / 334
页数:8
相关论文
共 50 条
  • [21] Towards Practical Protection of Data Access Pattern to Cloud Storage
    Ma, Qiumao
    Zhang, Wensheng
    2018 IEEE MILITARY COMMUNICATIONS CONFERENCE (MILCOM 2018), 2018, : 243 - 248
  • [22] Hierarchical Access Control with Scalable Data Sharing in Cloud Storage
    Qiu, Zhenyao
    Zhang, Zhiwei
    Tan, Shichong
    Wang, Jianfeng
    Tao, Xiaoling
    JOURNAL OF INTERNET TECHNOLOGY, 2019, 20 (03): : 663 - 676
  • [23] Combining Mobile and Cloud Storage for Providing Ubiquitous Data Access
    Soares, Joao
    Preguica, Nuno
    EURO-PAR 2011 PARALLEL PROCESSING, PT 1, 2011, 6852 : 516 - 527
  • [24] The Data Sharing Security System of Cloud Storage
    Liu, Shenglong
    Zhang, Ge
    Xia, Yuxiao
    Yang, Ruxia
    INTERNATIONAL SYMPOSIUM ON ARTIFICIAL INTELLIGENCE AND ROBOTICS 2021, 2021, 11884
  • [25] On construction of a distributed data storage system in cloud
    Chao-Tung Yang
    Wen-Chung Shih
    Chih-Lin Huang
    Fuu-Cheng Jiang
    William Cheng-Chung Chu
    Computing, 2016, 98 : 93 - 118
  • [26] Design of Cloud Data Storage and Processing System
    Zhou, Baoke
    Chou, Wusheng
    2018 INTERNATIONAL CONFERENCE ON BIG DATA AND ARTIFICIAL INTELLIGENCE (BDAI 2018), 2018, : 6 - 9
  • [27] Storage System for Point Cloud Geographic Data
    Svalec, Marian
    Takac, Lubos
    Zabovsky, Michal
    INTERNATIONAL CONFERENCE ON COMPUTER SCIENCE AND INFORMATION ENGINEERING (CSIE 2015), 2015, : 304 - 310
  • [28] Data Replica Placement in Cloud Storage System
    Zhang Tao
    PROCEEDINGS OF THE 1ST INTERNATIONAL WORKSHOP ON CLOUD COMPUTING AND INFORMATION SECURITY (CCIS 2013), 2013, 52 : 551 - 554
  • [29] On construction of a distributed data storage system in cloud
    Yang, Chao-Tung
    Shih, Wen-Chung
    Huang, Chih-Lin
    Jiang, Fuu-Cheng
    Chu, William Cheng-Chung
    COMPUTING, 2016, 98 (1-2) : 93 - 118
  • [30] Research on Data Block Storage Strategy in Cloud Storage System
    Guo, Yuanyuan
    Hao, Jianjun
    Guo, Yijun
    Luo, Tao
    PROCEEDINGS OF 2017 3RD IEEE INTERNATIONAL CONFERENCE ON COMPUTER AND COMMUNICATIONS (ICCC), 2017, : 2394 - 2397