About | Contact Us | Register | Login
ProceedingsSeriesJournalsSearchEAI
ew 26(1):

Editorial

Reinforcement Learning-enhanced Policy-aware Modeling of Smart Grid Efficiency under Carbon Constraints: Integration of SBM DEA and Dynamic Policy Response Simulation

Download1 download
Cite
BibTeX Plain Text
  • @ARTICLE{10.4108/ew.12473,
        author={Gemei Shi},
        title={Reinforcement Learning-enhanced Policy-aware Modeling of Smart Grid Efficiency under Carbon Constraints: Integration of SBM DEA and Dynamic Policy Response Simulation},
        journal={EAI Endorsed Transactions on Energy Web},
        volume={13},
        number={1},
        publisher={EAI},
        journal_a={EW},
        year={2026},
        month={6},
        keywords={artificial intelligence, reinforcement learning, smart grid, carbon reduction efficiency, SBM-DEA, dynamic policy response},
        doi={10.4108/ew.12473}
    }
    
  • Gemei Shi
    Year: 2026
    Reinforcement Learning-enhanced Policy-aware Modeling of Smart Grid Efficiency under Carbon Constraints: Integration of SBM DEA and Dynamic Policy Response Simulation
    EW
    EAI
    DOI: 10.4108/ew.12473
Gemei Shi1,*
  • 1: Guangzhou University of Software
*Contact email: 609084168@qq.com

Abstract

Artificial intelligence-based sensing, forecasting, and decision optimization are being rapidly integrated into smart grid operations. Although reinforcement learning has created new opportunities for dispatch optimization and low-carbon transition under carbon constraints, most existing studies focus primarily on short-term economic or operational objectives and rarely incorporate system-level carbon reduction efficiency benchmarks into the learning process. To address this gap, this study proposes an integrated SBM-DEA and reinforcement learning framework for policy-aware smart-grid dispatch, in which carbon reduction efficiency scores are transformed from static evaluation results into dynamic learning signals for dispatch optimization. Using panel data from 30 provinces in China over the period 2011–2022, this study develops an indicator system covering capital input, labor input, electricity service output, and electricity-related carbon dioxide emissions. An SBM-DEA model with undesirable outputs is employed to measure the carbon reduction efficiency of smart grids. The estimated efficiency scores are then embedded into both the state representation and reward function of a reinforcement learning framework, where the agent learns dispatch policies that balance economic performance, carbon constraints, and efficiency improvement. A dynamic policy response simulation environment is further constructed, incorporating a hybrid energy storage system comprising battery storage and pumped hydro storage. The results show that the carbon reduction efficiency of smart grids in China exhibits stage-specific fluctuations, with annual average values ranging from 0.505 to 0.568 and pronounced interprovincial disparities. In the simulation experiments, the reinforcement learning agent trained with efficiency-based penalties achieves 7.3% lower operational costs and 8.5% higher average efficiency compared to an economic-only agent. The trained policies also exhibit clear policy-responsive behavior: when carbon prices rise, hybrid storage utilization increases and coal-fired generation declines. The main innovation of this study is that it integrates historical efficiency benchmarking with reinforcement learning-based dispatch optimization, providing a policy-aware and efficiency-guided decision-support framework for carbon-constrained smart grids.

Keywords
artificial intelligence, reinforcement learning, smart grid, carbon reduction efficiency, SBM-DEA, dynamic policy response
Received
2026-04-03
Accepted
2026-05-25
Published
2026-06-04
Publisher
EAI
http://dx.doi.org/10.4108/ew.12473

Copyright © 2026 Geimei Shi et al., licensed to EAI. This is an open access article distributed under the terms of the CC BY-NC-SA 4.0, which permits copying, redistributing, remixing, transformation, and building upon the material in any medium so long as the original work is properly cited.

EBSCOProQuestDBLPDOAJPortico
EAI Logo

About EAI

  • Who We Are
  • Leadership
  • Research Areas
  • Partners
  • Media Center
  • Cookie Preferences

Community

  • Membership
  • Conference
  • Recognition
  • Sponsor Us

Publish with EAI

  • Publishing
  • Journals
  • Proceedings
  • Books
  • EUDL