Abstract
This article proposes a multiagent (MA) deep reinforcement learning (DRL) based autonomous input voltage sharing (IVS) control and triple phase shift modulation method for input-series output-parallel (ISOP) dual active bridge (DAB) converters to solve the three challenges: the uncertainties of the dc microgrid, the power balance problem, and the current stress minimization of the converter. Specifically, the control and modulation problem of the ISOP-DAB converter is formed as a Markov game with several DRL agents. Subsequently, the MA twin-delayed deep deterministic policy gradient (MA-TD3) algorithm is applied to train the DRL agents in an offline manner. After the training process, the multiple agents can provide online control decisions for the ISOP-DAB converter to balance the IVS, and minimize the current stress among different submodules. Without accurate model information, the proposed method can adaptively obtain the optimal modulation variable combinations in a stochastic and uncertain environment. Simulation and experimental results verify the effectiveness of the proposed MA-TD3-based algorithm. © 2022 IEEE.
| Original language | English |
|---|---|
| Pages (from-to) | 2985-3000 |
| Journal | IEEE Transactions on Power Electronics |
| Volume | 38 |
| Issue number | 3 |
| Online published | 2 Nov 2022 |
| DOIs | |
| Publication status | Published - Mar 2023 |
| Externally published | Yes |
Funding
This work was supported in part by the National Research Foundation (NRF) of Singapore, Rolls-Royce Singapore Pte., Ltd., and in part by Nanyang Technological University, Singapore.
Research Keywords
- input voltage sharing (IVS)
- Input-series output-parallel-connected dual active bridge (ISOP-DAB) converter
- multiagent twin-delayed deep deterministic policy gradient (MA-TD3)
- triple phase shift modulation
Fingerprint
Dive into the research topics of 'Autonomous Input Voltage Sharing Control and Triple Phase Shift Modulation Method for ISOP-DAB Converter in DC Microgrid: A Multiagent Deep Reinforcement Learning-Based Method'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver