Código QR

Clustering-based Failed goal Aware Hindsight Experience Replay

In a multi-goal reinforcement learning environment, an agent learns a policy to perform tasks with multiple goals from experiences gained through exploration. In environments with sparse binary rewards, the replay buffer contains few successful experiences, posing a challenge for sampling efficiency...

ver descrição completa

Na minha lista:
Detalhes bibliográficos
Principais autores: Taeyoung Kim, Taemin Kang, Haechan Jeong, Dongsoo Har
Formato: Artigo
Idioma:Inglês
Publicado em: PeerJ Inc. 2024-12-01
Colecção:PeerJ Computer Science
Assuntos:
Acesso em linha:https://peerj.com/articles/cs-2588.pdf
Tags: Adicionar Tag
Sem tags, seja o primeiro a adicionar uma tag!