Deep Recurrent Q-learning Method for Area Traffic Coordination Control
Saijiang Shi & Feng Chen · Journal of Advances in Mathematics and Computer Science · 2018
In order to improve the performance of Deep Q-learning when dealing with the area traffic control which is a partially observable Markov decision process. This paper introduces Deep Recurrent Q-learning by changing the fully connected network layers to LSTM layers. On the other h...
Open access
Research Article
10.9734/JAMCS/2018/41281