DRL Based Traffic Signal Control Method Featuring Masked Approach to Redress Transmission Error in ITS
摘要
Due to the everchanging dynamics of traffic situation, managing real-time traffic congestion with great efficiency is exceedingly challenging. Deep Reinforcement Learning (DRL) in Intelligent Transportation System (ITS) under the concept of Edge computing is an approach that determines the optimal traffic signal strategy for dealing with traffic congestion. Optimizing traffic signal with a DRL agent involves transmitting state information collected by edge devices. However, network congestion, device malfunctions, and transmission delays often impede the transmission of information. Consequently, the decision-making capacity of the agent suffers from inadequate information, leading to decreased efficacy. To mitigate this issue, the study proposes two distinct masking methods on input states. A single DRL agent deals with these masked inputs from the environment through the Edge devices. In order to train the agent, the DRL algorithm Proximal Policy Optimization (PPO) is implemented in five different neural network models including the state-of-the-art Transformer network which can accurately model spatial dependence and capture the persistence of sequential data. To validate the feasibility of the agent, simulation experiments are conducted in hypothetical road network and real-time road map. The experiments utilize waiting time, fuel consumption, and