QR-koda

Clustering-based Failed goal Aware Hindsight Experience Replay

In a multi-goal reinforcement learning environment, an agent learns a policy to perform tasks with multiple goals from experiences gained through exploration. In environments with sparse binary rewards, the replay buffer contains few successful experiences, posing a challenge for sampling efficiency...

Olles dieđut

Furkejuvvon:
Bibliográfalaš dieđut
Váldodahkkit: Taeyoung Kim, Taemin Kang, Haechan Jeong, Dongsoo Har
Materiálatiipa: Artigo
Giella:Inglês
Almmustuhtton: PeerJ Inc. 2024-12-01
Ráidu:PeerJ Computer Science
Fáttát:
Liŋkkat:https://peerj.com/articles/cs-2588.pdf
Fáddágilkorat: Lasit fáddágilkoriid
Eai fáddágilkorat, Lasit vuosttaš fáddágilkora!