QR Kod

Directly Attention loss adjusted prioritized experience replay

Abstract Prioritized Experience Replay enables the model to learn more about relatively important samples by artificially changing their accessed frequencies. However, this non-uniform sampling method shifts the state-action distribution that is originally used to estimate Q-value functions, which b...

Ful tanımlama

Kaydedildi:
Detaylı Bibliyografya
Asıl Yazarlar: Zhuoying Chen, Huiping Li, Zhaoxu Wang
Materyal Türü: Artigo
Dil:Inglês
Baskı/Yayın Bilgisi: Springer 2025-04-01
Seri Bilgileri:Complex & Intelligent Systems
Konular:
Online Erişim:https://doi.org/10.1007/s40747-025-01852-6
Etiketler: Etiketle
Etiket eklenmemiş, İlk siz ekleyin!