Skip to content
Open access

Safety-Constrained Deep Reinforcement Learning for Source–Load–Storage Coordinated Operation of Green Low-Carbon Data Centers

Jul 2026 · Energies · 0 citations · 32 references

TL;DR

A safety-constrained deep reinforcement learning framework for source–load–storage coordinated operation of a grid-connected green data center and shows how separating reward learning, cumulative safety pricing, and one-step engineering projection changes low-carbon dispatch within the specified model.

Abstract

Green low-carbon data centers operate as coupled cyber-energy systems whose dispatch must coordinate renewable generation, grid exchange, battery storage, cooling load, flexible computing workload, carbon-intensity signals, and reliability constraints. This study develops and evaluates a safety-constrained deep reinforcement learning framework for source–load–storage coordinated operation of a grid-connected green data center. The operating problem is formulated as a constrained Markov decision process with state variables describing the IT load, deferrable workload backlog, renewable availability, electricity price, marginal carbon intensity, battery state of charge, server-room temperature, reserve margin, and calendar context. The action space covers grid import and export, renewable utilization, storage charge and discharge, workload shifting, and cooling control. The learning architecture combines a constrained actor–critic policy, adaptive Lagrangian safety critics, and a control barrier function (CBF)-based action shield that projects unsafe actions onto an explicitly defined operating set before plant execution. The shield is specified as a low-dimensional quadratic projection over state-dependent SOC, thermal, reserve, SLA, and grid-interface constraints, while cumulative risks are priced through Lagrangian safety budgets during policy training. The evaluation uses a controlled and auditable benchmark simulation with normalized public-data-compatible profiles, declared scenarios, random seeds, neural-network settings, and mechanism-matched baselines; it is not a telemetry-based verification or hardware certification of a deployed data center. Within this declared benchmark, the proposed safe DRL controller produces a simulated 13.1% emission reduction relative to the Rule-based controller, 95.8% renewable utilization, a normalized annual cost of 0.91, and fewer boundary contacts than the tested unconstrained, Lagrangian-only, and shield-only PPO variants. These percentages are simulator outputs relative to the stated benchmark and must not be interpreted as measured field savings. The results show how separating reward learning, cumulative safety pricing, and one-step engineering projection changes low-carbon dispatch within the specified model.

Read PDF

Similar papers

Open access Aug 2026

Grid-Interactive Hyperscale Data Centers: Deep Reinforcement Learning for Joint Workload-Cooling Scheduling to Enable Demand Response and Renewable Integration

A deep reinforcement learning (DRL) framework that jointly co-schedules computing and thermal resources so that a hyperscale data center can operate as a grid-interactive flexible load and supports the evolution of hyperscale data centers from passive electricity consumers toward active, grid-interactive participants i...

You-Chun Qiu · 0 citations
Open access Jul 2026

Physics-Informed Distributionally Robust Multi-Agent Reinforcement Learning for Coordinated New-Type Power System Operation

High renewable penetration and large-scale green hydrogen production are accelerating the formation of the new-type power system (NTPS), in which electrical dispatch, electrolysis, hydrogen storage, fuel-cell reconversion, and flexible demand must be coordinated under nonlinear network physics and uncertain renewable,...

Fei Liu, Outing Zhang, Jun Yin et al. · 0 citations
Open access Aug 2026

Multi-Objective Optimization for Data Center HVAC Systems Based on Edge–Cloud Collaborative Deep Reinforcement Learning

This paper proposes an edge-cloud collaborative physics-informed reinforcement learning framework for production data center HVAC control that integrates a physics-informed cold-start solution using Adaptive Particle Swarm Optimization, a three-time-scale edge–cloud architecture, and a constraint-aware safe projection...

Shichao Huang, Yi-Bing Zhou, Yuan Liu · 0 citations
Jul 2026

Safe Multi-Agent Collaborative Learning for Networked Grid Operation Under Power Network Coupling Constraints

This paper formulate networked grid operation as a constrained decentralized partially observable Markov decision process and proposes a safe multi-agent collaborative learning framework that aims to reduce operating cost, load shedding, renewable curtailment, and carbon-relevant corrective burden.

Jia-Yi Zhang, Bing Fang, Huan-Xiu Xiao et al. · 0 citations
Open access 2026

Deep Reinforcement Learning-Based Adaptive Switching for Risk-Cost-Optimized Renewable Smart Grids

A deep reinforcement learning-based adaptive switching framework to enhance renewable utilization while minimizing operational risk and economic cost is proposed and provides a scalable and intelligent solution for industrial smart grid applications.

M. Meyyappan, P. Avirajamanjula, P. Marimuthu et al. · 0 citations
Open access Sep 2026

Machine-Learning-Assisted Multi-Energy Coupling and Battery–Grid Coordination for Deep Decarbonization of Smart Integrated Energy Systems: Modeling, Optimization, and Applications

A machine-learning-assisted, renewable-driven framework for multi-energy coupling and scenario-based multi-objective optimization of electricity–heat–hydrogen–storage systems and provides a data-driven modeling and decision framework for battery–grid coordination and deep decarbonization in smart integrated energy syst...

Yao Tong, Hai-Ling Ma, Fu-Yi Du · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.