Decoupling Datacenter Storage Studies from Access to Large-Scale Applications

被引:1
作者
Delimitrou, Christina [1 ]
Sankar, Sriram [2 ]
Vaid, Kushagra [2 ]
Kozyrakis, Christos [1 ]
机构
[1] Stanford Univ, Stanford, CA 94305 USA
[2] Microsoft Corp, Seattle, WA USA
关键词
Modeling of computer architecture; Super (very large) computers; Mass storage; Modeling techniques;
D O I
10.1109/L-CA.2011.37
中图分类号
TP3 [计算技术、计算机技术];
学科分类号
0812 ;
摘要
Suboptimal storage design has significant cost and power impact in large-scale datacenters (DCs). Performance, power and cost-optimized systems require deep understanding of target workloads, and mechanisms to effectively model different storage design choices. Traditional benchmarking is invalid in cloud data-stores, representative storage profiles are hard to obtain, while replaying applications in different storage configurations is impractical both in cost and time. Despite these issues, current workload generators are not able to reproduce key aspects of real application patterns (e.g., spatial/temporal locality, I/O intensity). In this paper, we propose a modeling and generation framework for large-scale storage applications. As part of this framework we use a state diagram-based storage model, extend it to a hierarchical representation, and implement a tool that consistently recreates DC application I/O loads. We present the principal features of the framework that allow accurate modeling and generation of storage workloads, and the validation process performed against ten original DC application traces. Finally, we explore two practical applications of this methodology: SSD caching and defragmentation benefits on enterprise storage. Since knowledge of the workload's spatial and temporal locality is necessary to model these use cases, our framework was instrumental in quantifying their performance benefits. The proposed methodology provides detailed understanding of the storage activity of large-scale applications, and enables a wide spectrum of storage studies, without the requirement to access application code and full application deployment.
引用
收藏
页码:53 / 56
页数:4
相关论文
共 7 条
  • [1] Adaptec MaxIQ, 32GB SSD CACH PERF K
  • [2] Ahmad I., 2007, P IEEE IISWC BOST MA
  • [3] Delimitrou C., 2011, P EXERT CA MARCH
  • [4] SERVER ENGINEERING INSIGHTS FOR LARGE-SCALE ONLINE SERVICES
    Kozyrakis, Christos
    Kansal, Aman
    Sankar, Sriram
    Vaid, Kushagra
    [J]. IEEE MICRO, 2010, 30 (04) : 8 - 19
  • [5] Narayanan D., 2009, P EUROSYS NUR
  • [6] Sankar S., 2009, P IEEE IISWC TX
  • [7] Sankar S., 2010, P 1 WOSP SIPEW SAN J