QRコード

Hierarchical Reinforcement Learning Method Based on Trajectory Information

The option-based hierarchical reinforcement learning(O-HRL) algorithm has the characteristics of temporal abstraction,which can effectively deal with complex problems such as long-term temporal order and sparse rewards that are difficult to solve in reinforcement learning.The existing studies of O-H...

詳細記述

保存先:
書誌詳細
第一著者: XU Yapeng, LIU Quan, LI Junwei
フォーマット: Artigo
言語:Chinês
出版事項: Editorial office of Computer Science 2023-12-01
シリーズ:Jisuanji kexue
主題:
オンライン・アクセス:https://www.jsjkx.com/fileup/1002-137X/PDF/1002-137X-2023-50-12-314.pdf
タグ: タグ追加
タグなし, このレコードへの初めてのタグを付けませんか!