Clustering-based Failed goal Aware Hindsight Experience Replay
In a multi-goal reinforcement learning environment, an agent learns a policy to perform tasks with multiple goals from experiences gained through exploration. In environments with sparse binary rewards, the replay buffer contains few successful experiences, posing a challenge for sampling efficiency...
Furkejuvvon:
| Váldodahkkit: | , , , |
|---|---|
| Materiálatiipa: | Artigo |
| Giella: | Inglês |
| Almmustuhtton: |
PeerJ Inc.
2024-12-01
|
| Ráidu: | PeerJ Computer Science |
| Fáttát: | |
| Liŋkkat: | https://peerj.com/articles/cs-2588.pdf |
| Fáddágilkorat: |
Eai fáddágilkorat, Lasit vuosttaš fáddágilkora!
|
