详细信息

Model-free optimal consensus control for multi-agent systems via DDPG-based event-triggered adaptive dynamic programming method    

文献类型:期刊文献

英文题名:Model-free optimal consensus control for multi-agent systems via DDPG-based event-triggered adaptive dynamic programming method

作者:Zhu, Pengfei[1];Wang, Xiaolin[1];Li, Fangfei[1,2];Qian, Siyu[1];Li, Haitao[3]

机构:[1]East China Univ Sci & Technol, Sch Math, Shanghai, Peoples R China;[2]East China Univ Sci & Technol, Key Lab Smart Mfg Energy Chem Proc, Minist Educ, Shanghai, Peoples R China;[3]Shandong Normal Univ, Sch Math & Stat, Jinan 250014, Peoples R China

年份:2026

卷号:6

期号:1

起止页码:97

外文期刊名:MATHEMATICAL MODELLING AND CONTROL

收录:WOS:【ESCI(收录号:WOS:001711432900004)】;

语种:英文

外文关键词:augmented system; event-triggered adaptive dynamic programming; deep deterministic policy gradient; event-triggered condition

摘要:This paper primarily addresses the design of distributed optimal cooperative controllers and the utilization of a reinforcement learning (RL)-based event-triggered mechanism for multi-agent systems (MASs) with unknown dynamics. By setting an extra compensator, the augmented system is constructed to overcome the dependence for system dynamics. Then, to address the issue of computational burden, we utilize an event-triggered mechanism based on reinforcement learning (RL) and neural networks (NNs) to implement the adaptive dynamic programming (ADP) algorithm. Additionally, we take into consideration the trade-off between computational burden and achieving consensus control by introducing a weighting factor in the reward design for MASs. With this reward design, we present an algorithm based on the deep deterministic policy gradient (DDPG) algorithm to learn the event-triggered condition for MASs and achieve a balance between these two factors. The event-triggered mechanism of our algorithm can also identify constraints such as time limitations or computational resource restrictions, aiming to achieve consensus control without violating these constraints. We demonstrate the absence of Zeno behavior and the uniform ultimate boundedness (UUB) of both local consensus error and weight estimation error. Finally, simulation results illustrate the effectiveness of the control algorithm and the weighting factor.

参考文献:

正在载入数据...

版权所有©华东理工大学 重庆维普资讯有限公司 渝B2-20050021-7 
渝公网安备 50019002500408号 违法和不良信息举报中心