Hierarchical Reinforcement Learning Method Based on Trajectory Information
The option-based hierarchical reinforcement learning(O-HRL) algorithm has the characteristics of temporal abstraction,which can effectively deal with complex problems such as long-term temporal order and sparse rewards that are difficult to solve in reinforcement learning.The existing studies of O-H...
保存先:
| 第一著者: | |
|---|---|
| フォーマット: | Artigo |
| 言語: | Chinês |
| 出版事項: |
Editorial office of Computer Science
2023-12-01
|
| シリーズ: | Jisuanji kexue |
| 主題: | |
| オンライン・アクセス: | https://www.jsjkx.com/fileup/1002-137X/PDF/1002-137X-2023-50-12-314.pdf |
| タグ: |
タグなし, このレコードへの初めてのタグを付けませんか!
|
