Directly Attention loss adjusted prioritized experience replay
Abstract Prioritized Experience Replay enables the model to learn more about relatively important samples by artificially changing their accessed frequencies. However, this non-uniform sampling method shifts the state-action distribution that is originally used to estimate Q-value functions, which b...
Kaydedildi:
| Asıl Yazarlar: | , , |
|---|---|
| Materyal Türü: | Artigo |
| Dil: | Inglês |
| Baskı/Yayın Bilgisi: |
Springer
2025-04-01
|
| Seri Bilgileri: | Complex & Intelligent Systems |
| Konular: | |
| Online Erişim: | https://doi.org/10.1007/s40747-025-01852-6 |
| Etiketler: |
Etiket eklenmemiş, İlk siz ekleyin!
|
