A user-centric, black-box performance characterization of ESSDs from Amazon AWS and Alibaba Cloud is conducted and it is hoped these contributions can serve as a practical reference for EBS users to understand and exploit the distinctive performance properties of ESSDs.
Abstract
Elastic block storage (EBS) with the storage-compute disaggregated architecture is a key component in modern cloud infrastructure. EBS offers users storage resources in the form of elastic solid-state drives (ESSDs). Nonetheless, despite recent efforts that have documented EBS architectures from the provider's perspective, how ESSDs perform differently from local SSDs and how host software should adapt accordingly have not been sufficiently studied. In this paper, we conduct a user-centric, black-box performance characterization of ESSDs from Amazon AWS and Alibaba Cloud. We make three main contributions: (1) an ESSD contract that presents four behavioral observations and five actionable implications for software adaptation, (2) a refined I/O rate-limiting model combining bandwidth-IOPS dual limiting and fine-grained token refilling to suppress latency spikes, and (3) a case study on RocksDB that derives four guidelines on cache management, I/O regulation, storage budget utilization, and compression algorithms. Collectively, we hope these contributions can serve as a practical reference for EBS users to understand and exploit the distinctive performance properties of ESSDs.
With the widespread adoption of cloud computing and virtualization, multi-tenant architecture has become the mainstream deployment model for data center storage. Solid-state drives (SSDs), leveraging high IOPS, low latency, and high parallelism, serve as the core storage medium. Nevertheless, internal resource contenti...
Dan-Dan Su, Qi-Hao Liu, Xin-Ming Li et al.· International Conference on...· 0 citations
This paper presents an in-depth analysis of data migration behavior in commercial HSSs, uncovering substantial performance variability when multiple migration tasks execute concurrently, and proposes PASCAL, a system-level bandwidth orchestration framework that improves performance robustness in production-grade HSSs.
Ji Zhang, Li Liu, André Brinkmann et al.· Proceedings of the VLDB Endo...· 0 citations
Large-scale online services—including web search, recommendation, and LLM inference workloads such as Retrieval-Augmented Generation (RAG) and KV-cache offloading—demand storage that handles petabyte-scale data under millisecond tail-latency SLAs. In-memory stores are cost-prohibitive at scale; disk-based systems sacri...
Ying-Xin Li, Kai Liu, Hanglun Xie· Proceedings of the VLDB Endo...· 0 citations
KV-cache offload is widely used to stretch GPU memory for LLM serving, but its storage behavior has not been characterized at the block-device level. In this paper, we study LMCache through realworld multi-session workloads that span same/different context $\times$ same/different prompt, using over 100 stateless reques...
Ying He, Dingsen Shi, Yanbo Dai et al.· 2026 International Conferenc...· 0 citations
: SystemC transaction-level modeling (TLM) is widely used for system integration, virtual platform simulation, and early software bring-up, yet storage and analysis of TLM traces remains unstandardized and largely driven by convenience or legacy tooling. This paper presents a comparative empirical evaluation of storage...
George Frazier, Dusti Johnson· International Conference on...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.