QR code

Directly Attention loss adjusted prioritized experience replay

Abstract Prioritized Experience Replay enables the model to learn more about relatively important samples by artificially changing their accessed frequencies. However, this non-uniform sampling method shifts the state-action distribution that is originally used to estimate Q-value functions, which b...

Volledige beschrijving

Bewaard in:
Bibliografische gegevens
Hoofdauteurs: Zhuoying Chen, Huiping Li, Zhaoxu Wang
Formaat: Artigo
Taal:Inglês
Gepubliceerd in: Springer 2025-04-01
Reeks:Complex & Intelligent Systems
Onderwerpen:
Online toegang:https://doi.org/10.1007/s40747-025-01852-6
Tags: Voeg label toe
Geen labels, Wees de eerste die dit record labelt!