{"title":"Heuristic dense reward shaping for learning-based map-free navigation of industrial automatic mobile robots.","authors":"Yizhi Wang, Yongfang Xie, Degang Xu, Jiahui Shi, Shiyu Fang, Weihua Gui","doi":"10.1016/j.isatra.2024.10.026","DOIUrl":null,"url":null,"abstract":"<p><p>This paper presents a map-free navigation approach for industrial automatic mobile robots (AMRs), designed to ensure computational efficiency, cost-effectiveness, and adaptability. Utilizing deep reinforcement learning (DRL), the system enables real-time decision-making without fixed markers or frequent map updates. The central contribution is the Heuristic Dense Reward Shaping (HDRS), inspired by potential field methods, which integrates domain knowledge to improve learning efficiency and minimize suboptimal actions. To address the simulation-to-reality gap, data augmentation with controlled sensor noise is applied during training, ensuring robustness and generalization for real-world deployment without fine-tuning. Training results underscore HDRS's superior convergence speed, training stability, and policy learning efficiency compared to baselines. Simulation and real-world evaluations establish HDRS-DRL as a competitive alternative, outperforming traditional approaches, and offering practical applicability in industrial settings.</p>","PeriodicalId":94059,"journal":{"name":"ISA transactions","volume":" ","pages":""},"PeriodicalIF":0.0000,"publicationDate":"2024-11-06","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"ISA transactions","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1016/j.isatra.2024.10.026","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}
引用次数: 0
Abstract
This paper presents a map-free navigation approach for industrial automatic mobile robots (AMRs), designed to ensure computational efficiency, cost-effectiveness, and adaptability. Utilizing deep reinforcement learning (DRL), the system enables real-time decision-making without fixed markers or frequent map updates. The central contribution is the Heuristic Dense Reward Shaping (HDRS), inspired by potential field methods, which integrates domain knowledge to improve learning efficiency and minimize suboptimal actions. To address the simulation-to-reality gap, data augmentation with controlled sensor noise is applied during training, ensuring robustness and generalization for real-world deployment without fine-tuning. Training results underscore HDRS's superior convergence speed, training stability, and policy learning efficiency compared to baselines. Simulation and real-world evaluations establish HDRS-DRL as a competitive alternative, outperforming traditional approaches, and offering practical applicability in industrial settings.