Mã QR

Human skill knowledge guided global trajectory policy reinforcement learning method

Traditional trajectory learning methods based on Imitation Learning (IL) only learn the existing trajectory knowledge from human demonstration. In this way, it can not adapt the trajectory knowledge to the task environment by interacting with the environment and fine-tuning the policy. To address th...

Mô tả đầy đủ

Đã lưu trong:
Chi tiết về thư mục
Những tác giả chính: Yajing Zang, Pengfei Wang, Fusheng Zha, Wei Guo, Chuanfeng Li, Lining Sun
Định dạng: Artigo
Ngôn ngữ:Inglês
Được phát hành: Frontiers Media S.A. 2024-03-01
Loạt:Frontiers in Neurorobotics
Những chủ đề:
Truy cập trực tuyến:https://www.frontiersin.org/articles/10.3389/fnbot.2024.1368243/full
Các nhãn: Thêm thẻ
Không có thẻ, Là người đầu tiên thẻ bản ghi này!