详细信息

Demand-Responsive Transport Dynamic Scheduling Optimization Based on Multi-agent Reinforcement Learning Under Mixed Demand  ( CPCI-S收录)  

文献类型:会议论文

英文题名:Demand-Responsive Transport Dynamic Scheduling Optimization Based on Multi-agent Reinforcement Learning Under Mixed Demand

作者:Wang, Jianrui[1];Li, Yi[1];Sun, Qiyu[1];Tang, Yang[1]

机构:[1]East China Univ Sci & Technol, Shanghai 200237, Peoples R China

会议论文集:33rd International Conference on Artificial Neural Networks and Machine Learning (ICANN)

会议日期:SEP 17-20, 2024

会议地点:Univ Italian Switzerland, Lugano, SWITZERLAND

主办单位:Univ Italian Switzerland

语种:英文

外文关键词:multi-agent reinforcement learning; demand-responsive transport; dynamic scheduling; route optimization

摘要:Demand-Responsive Transport (DRT) is an innovative mode of public transportation that focuses on individual passenger needs by offering customized transportation solutions. Most prior researches rely on historical passenger flow to generate static schemes and lack the optimization of optional dynamic demands from the perspectives of passengers and transportation agencies simultaneously. Therefore, this paper addresses the dynamic scheduling optimization problem of DRT under mixed demand, minimizing overall system costs and ensuring equitable passenger waiting times. We initially construct a dual-objective optimization model for DRT dynamic scheduling to solve this. Subsequently, we propose the Action-Refinement Multi-Agent Dueling Double Deep Q-Network (AR-MAD3QN) algorithm to tackle the challenge of simultaneous route optimization for a fleet of vehicles considering static and optional dynamic passenger demands under dynamic road conditions. Additionally, the action-refinement module improves the network structure of MAD3QN, preventing the generation of invalid and unstable actions and improving training efficiency. Experiments are conducted on the Sioux Falls network, with the AR-MAD3QN algorithm compared against baseline algorithms in different settings. The results show that our AR-MAD3QN algorithm exhibits superior optimization with faster and more stable convergence.

参考文献:

正在载入数据...

版权所有©华东理工大学 重庆维普资讯有限公司 渝B2-20050021-7 
渝公网安备 50019002500408号 违法和不良信息举报中心