{"title":"点对点赛车:演化与时间差异学习的初步研究","authors":"S. Lucas, J. Togelius","doi":"10.1109/CIG.2007.368107","DOIUrl":null,"url":null,"abstract":"This paper considers variations on an extremely simple form of car racing, the challenge being to visit as many way-points as possible in a fixed amount of time. The simplicity of the models enables a very thorough evaluation of various learning algorithms and control architectures, and enables other researchers to work on the same models with relative ease. The models are used to compare the performance of various hand-programmed controllers, and neural networks trained using evolution, and using temporal difference learning. Comparisons are also made between state-based and action-based controller architectures. The best controllers were obtained using evolution to learn the weights of state-evaluation neural networks, and these were greatly superior to human drivers","PeriodicalId":365269,"journal":{"name":"2007 IEEE Symposium on Computational Intelligence and Games","volume":"70 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2007-04-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"37","resultStr":"{\"title\":\"Point-to-Point Car Racing: an Initial Study of Evolution Versus Temporal Difference Learning\",\"authors\":\"S. Lucas, J. Togelius\",\"doi\":\"10.1109/CIG.2007.368107\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"This paper considers variations on an extremely simple form of car racing, the challenge being to visit as many way-points as possible in a fixed amount of time. The simplicity of the models enables a very thorough evaluation of various learning algorithms and control architectures, and enables other researchers to work on the same models with relative ease. The models are used to compare the performance of various hand-programmed controllers, and neural networks trained using evolution, and using temporal difference learning. Comparisons are also made between state-based and action-based controller architectures. The best controllers were obtained using evolution to learn the weights of state-evaluation neural networks, and these were greatly superior to human drivers\",\"PeriodicalId\":365269,\"journal\":{\"name\":\"2007 IEEE Symposium on Computational Intelligence and Games\",\"volume\":\"70 1\",\"pages\":\"0\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"2007-04-01\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"37\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"2007 IEEE Symposium on Computational Intelligence and Games\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.1109/CIG.2007.368107\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"2007 IEEE Symposium on Computational Intelligence and Games","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1109/CIG.2007.368107","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}
Point-to-Point Car Racing: an Initial Study of Evolution Versus Temporal Difference Learning
This paper considers variations on an extremely simple form of car racing, the challenge being to visit as many way-points as possible in a fixed amount of time. The simplicity of the models enables a very thorough evaluation of various learning algorithms and control architectures, and enables other researchers to work on the same models with relative ease. The models are used to compare the performance of various hand-programmed controllers, and neural networks trained using evolution, and using temporal difference learning. Comparisons are also made between state-based and action-based controller architectures. The best controllers were obtained using evolution to learn the weights of state-evaluation neural networks, and these were greatly superior to human drivers