Skip to content
Open access

Scalable Caching with Amazon ElastiCache Redis Cluster Mode: A Quantitative Performance Study

Jul 2026 · International journal of computer information systems and industrial management applications · 0 citations

Abstract

Enterprise applications increasingly depend on distributed caching to sustain sub-millisecond response times at scale. Amazon ElastiCache Redis, operating in cluster mode, provides horizontal partitioning across configurable shard topologies, enabling throughput and memory capacity to grow in proportion to demand. While many organizations have adopted cluster configurations, empirical guidance on topology selection, key distribution optimization, and the measurable performance impact of individual tuning techniques remains sparse. This paper addresses that gap through systematic benchmarking across multiple cluster topologies (3 to 90 shards), three Graviton-based instance families (m6g, r6g, r7g), three workload profiles, and five optimization techniques, augmented by client library analysis, memory optimization guidance, and production cost validation. Production case studies from financial services, e-commerce, and real-time analytics platforms validate laboratory findings. Results offer empirical guidance for cloud architects designing caching architectures that balance latency requirements, horizontal scalability objectives, and infrastructure cost efficiency.

Read PDF

Similar papers

Preprint Aug 2026

Performance and Cost-Aware Cache Provisioning

This paper presents a novel hybrid segmented policy that reduces capacity requirements while keeping processing costs low and shows that dynamically adjusting the segment ratio in segmented policies based on historical workload patterns enhances efficiency.

R. Tanvir, G. Kesidis · 0 citations
Book Open access Sep 2026

RDPart: A reuse-based OS-level cache-partitioning policy for fairness optimization in cloud data centers

In cloud data centers, the colocation of multiple services and applications on the same physical server is crucial for maximizing resource utilization and reducing utility costs. Unfortunately, contention for shared resources across cores, such as the Last-Level Cache (LLC), may lead to severe interference among applic...

Javier Aznal, J. C. Saez, Carlos Bilbao · 1 citation
Conference Aug 2026

QoS Guarantees and Performance Isolation for Multi-tenant SSDs in Cloud Environments

With the widespread adoption of cloud computing and virtualization, multi-tenant architecture has become the mainstream deployment model for data center storage. Solid-state drives (SSDs), leveraging high IOPS, low latency, and high parallelism, serve as the core storage medium. Nevertheless, internal resource contenti...

Dan-Dan Su, Qi-Hao Liu, Xin-Ming Li et al. · 0 citations
Open access Aug 2026

EMC+: An Opportunistic Elasticity Method for Improving System Throughput and CPU Utilization in Cloud Data Centers

The new EMC+ proposal is an OS‐driven elasticity manager for container‐based environments that continuously estimates idle core cycles left by regular (inelastic) applications, and reallocates idle cores to elastic ones, even during short time intervals, and has minimal impact on the performance and QoS of colocated in...

J. C. Saez, Carlos Bilbao, Manuel Prieto-Matías · 0 citations

Towards Practical Latency SLOs on Cloud Data Warehouses

This work outlines AutoSLO, a latency-SLO-aware work-load management framework for multi-cluster cloud data ware-houses that includes a periodic Policy Tuner that proactively plans resources using workload fore-casts, an SLO-aware reactive Autoscaler that adjusts the active cluster set based on the observed workload, a...

Markos Markakis, T. Kraska · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.