This work uses Apache Spark running on Kubernetes as a case-study dataflow runtime and cluster resource manager to compare model-based energy estimates to Intel RAPL package and DRAM energy on an AWS bare-metal cloud and an on-premises cluster, comparing different CPU usage signals and memory coefficients.
Abstract
Distributed batch data processing applications are widely executed on cloud-based resources where restricted user access to node-level hardware energy counters hinders transparent sustainability accounting. Energy and carbon attribution methodologies therefore depend on power models and available resource utilisation traces, yet the accuracy of these estimates has to be validated while direct counters are available. In this work, we use Apache Spark running on Kubernetes as a case-study dataflow runtime and cluster resource manager to compare model-based energy estimates to Intel RAPL package and DRAM energy on an AWS bare-metal cloud and an on-premises cluster, comparing different CPU usage signals and memory coefficients. We show that external monitoring improves signed package-energy error relative to Spark task traces, reducing underestimation from -29.58% to -24.41% on AWS and from -24.00% to -16.22% on-premises.
Build pipelines are integral to modern software development, yet their energy footprint remains largely invisible to practitioners. Existing CI energy tools either rely on model-based estimation (due to hardware access restrictions in cloud runners) or report only total pipeline energy without decomposing it into meani...
Jérémy Woirhaye, François Gibier, Romain Rouvoy· 0 citations
Results show that the proposed framework for thermal-aware and carbon-efficient workload allocation could effectively select out the most suitable servers for workloads not only from viewpoint of carbon usage but also from the thermal perspectives, and at the same time, it also reduces the amount of cooling required an...
Chandan Hegde, Adarsh Bilimisi, Pruthvik J· International Journal of Lat...· 0 citations
Aethon is presented, a memory offloading system for public clouds that maximizes offloaded data while ensuring each application satisfies the service-level agreement (SLA).
Guang-Qiang Luan, Pu Pang, Quan Chen et al.· Proceedings of the Internati...· 0 citations
A carbon-aware routing framework that distributes function-calling queries across a three-tier edge-cloud architecture, combining edge and cloud LLMs on heterogeneous hardware and matches cloud-level accuracy while reducing operational carbon emissions by $4\times on average.
Aikaterini Maria Panteleaki, Varatheepan Paramanayakam, S. Tragoudas et al.· 1 citation
The energy consumption of digital infrastructures and applications constitutes a critical global concern. To build a sustainable digital future, software engineering must prioritize energy efficiency alongside traditional performance metrics. Actors provide an attractive model for distributed large-scale applications;...
I. Samus, Mario Südholt, C. De Roover et al.· Workshop Proceedings of the...· 0 citations
Cloud computing and high-performance computing (HPC) typically follow different paradigms: cloud services are often orchestrated using Kubernetes, whereas HPC workloads are managed through batch schedulers such as Slurm. Growing demand for shared computational resources increases the need for interoperability between t...
M. Mačernis· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.