Human skill knowledge guided global trajectory policy reinforcement learning method
Traditional trajectory learning methods based on Imitation Learning (IL) only learn the existing trajectory knowledge from human demonstration. In this way, it can not adapt the trajectory knowledge to the task environment by interacting with the environment and fine-tuning the policy. To address th...
Đã lưu trong:
| Những tác giả chính: | , , , , , |
|---|---|
| Định dạng: | Artigo |
| Ngôn ngữ: | Inglês |
| Được phát hành: |
Frontiers Media S.A.
2024-03-01
|
| Loạt: | Frontiers in Neurorobotics |
| Những chủ đề: | |
| Truy cập trực tuyến: | https://www.frontiersin.org/articles/10.3389/fnbot.2024.1368243/full |
| Các nhãn: |
Không có thẻ, Là người đầu tiên thẻ bản ghi này!
|
