Application of reinforcement learning with continuous state space to ramp metering in real-world conditions

K. Rezaee, B. Abdulhai, H. Abdelgawad
{"title":"Application of reinforcement learning with continuous state space to ramp metering in real-world conditions","authors":"K. Rezaee, B. Abdulhai, H. Abdelgawad","doi":"10.1109/ITSC.2012.6338837","DOIUrl":null,"url":null,"abstract":"In this paper we introduce a new approach to Freeway Ramp Metering (RM) based on Reinforcement Learning (RL) with focus on real-life experiments in a case study in the City of Toronto. Typical RL methods consider discrete state representation that lead to slow convergence in complex problems. Continuous representation of state space has the potential to significantly improve the learning speed and therefore enables tackling large-scale complex problems. A robust approach based on local regression, named k nearest neighbors temporal difference (kNN-TD), is employed to represent state space continuously in the RL environment. The performance of the new algorithm is compared against the ALINEA controller and typical RL methods using a micro-simulation testbed in Paramics. The results show that RM using the kNN-TD method can reduce total network travel time by 44% compared to the do-nothing case (without RM) and by 17% compared to ALINEA.","PeriodicalId":184458,"journal":{"name":"International Conference on Intelligent Transportation Systems","volume":"144 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2012-10-25","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"28","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"International Conference on Intelligent Transportation Systems","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1109/ITSC.2012.6338837","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}
引用次数: 28

Abstract

In this paper we introduce a new approach to Freeway Ramp Metering (RM) based on Reinforcement Learning (RL) with focus on real-life experiments in a case study in the City of Toronto. Typical RL methods consider discrete state representation that lead to slow convergence in complex problems. Continuous representation of state space has the potential to significantly improve the learning speed and therefore enables tackling large-scale complex problems. A robust approach based on local regression, named k nearest neighbors temporal difference (kNN-TD), is employed to represent state space continuously in the RL environment. The performance of the new algorithm is compared against the ALINEA controller and typical RL methods using a micro-simulation testbed in Paramics. The results show that RM using the kNN-TD method can reduce total network travel time by 44% compared to the do-nothing case (without RM) and by 17% compared to ALINEA.
查看原文
分享 分享
微信好友 朋友圈 QQ好友 复制链接
本刊更多论文
连续状态空间强化学习在匝道测量中的应用
在本文中,我们介绍了一种基于强化学习(RL)的高速公路匝道计量(RM)的新方法,并以多伦多市为例进行了实际实验研究。典型的强化学习方法考虑离散状态表示,导致复杂问题的缓慢收敛。状态空间的连续表示具有显著提高学习速度的潜力,因此能够解决大规模的复杂问题。采用一种基于局部回归的鲁棒方法k近邻时间差分(kNN-TD)来连续表示RL环境下的状态空间。在Paramics的微仿真试验台上,将新算法的性能与ALINEA控制器和典型RL方法进行了比较。结果表明,使用kNN-TD方法的RM与不做任何事情(不做RM)相比可减少44%的总网络旅行时间,与ALINEA相比可减少17%的总网络旅行时间。
本文章由计算机程序翻译,如有差异,请以英文原文为准。
求助全文
约1分钟内获得全文 去求助
来源期刊
自引率
0.00%
发文量
0
期刊最新文献
Combining K-means method and complex network analysis to evaluate city mobility Goal-Driven Context-Aware Data Filtering in IoT-Based Systems Vision-Based Driver Assistance Systems: Survey, Taxonomy and Advances An Improved FastSLAM Algorithm for Autonomous Vehicle Based on the Strong Tracking Square Root Central Difference Kalman Filter Planning of High-Level Maneuver Sequences on Semantic State Spaces
×
引用
GB/T 7714-2015
复制
MLA
复制
APA
复制
导出至
BibTeX EndNote RefMan NoteFirst NoteExpress
×
×
提示
您的信息不完整,为了账户安全,请先补充。
现在去补充
×
提示
您因"违规操作"
具体请查看互助需知
我知道了
×
提示
现在去查看 取消
×
提示
确定
0
微信
客服QQ
Book学术公众号 扫码关注我们
反馈
×
意见反馈
请填写您的意见或建议
请填写您的手机或邮箱
已复制链接
已复制链接
快去分享给好友吧!
我知道了
×
扫码分享
扫码分享
Book学术官方微信
Book学术文献互助
Book学术文献互助群
群 号:481959085
Book学术
文献互助 智能选刊 最新文献 互助须知 联系我们:info@booksci.cn
Book学术提供免费学术资源搜索服务,方便国内外学者检索中英文文献。致力于提供最便捷和优质的服务体验。
Copyright © 2023 Book学术 All rights reserved.
ghs 京公网安备 11010802042870号 京ICP备2023020795号-1