{"title":"Performance Evaluation of Tile Coding in Reinforcement Learning","authors":"Kenji Ota, T. Ozeki","doi":"10.1145/2814940.2814975","DOIUrl":null,"url":null,"abstract":"Reinforcement learning is one of research fields in artificial intelligence. The learning method usually assumes a discrete state in computer simulations. However, we must treat a continuous value in a realistic situation. In this paper, we investigate various techniques of the tile cording scheme which is a representative technique to handle continuous states. We check the performance of single tiling, multiple tiling, time-shift method and the proposed method in the issue of space search.","PeriodicalId":427567,"journal":{"name":"Proceedings of the 3rd International Conference on Human-Agent Interaction","volume":"48 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2015-10-21","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"1","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Proceedings of the 3rd International Conference on Human-Agent Interaction","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1145/2814940.2814975","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}
引用次数: 1
Abstract
Reinforcement learning is one of research fields in artificial intelligence. The learning method usually assumes a discrete state in computer simulations. However, we must treat a continuous value in a realistic situation. In this paper, we investigate various techniques of the tile cording scheme which is a representative technique to handle continuous states. We check the performance of single tiling, multiple tiling, time-shift method and the proposed method in the issue of space search.