Comparative Analysis of A3C and PPO Algorithms in Reinforcement Learning: A Survey on General Environments
This research article presents a comparison between two mainstream Deep Reinforcement Learning (DRL) algorithms, Asynchronous Advantage Actor-Critic (A3C) and Proximal Policy Optimization (PPO), in the context of two diverse environments: CartPole and Lunar Lander. DRL algorithms are widely known fo...
Kaydedildi:
| Asıl Yazarlar: | , , |
|---|---|
| Materyal Türü: | Artigo |
| Dil: | Inglês |
| Baskı/Yayın Bilgisi: |
IEEE
2024-01-01
|
| Seri Bilgileri: | IEEE Access |
| Konular: | |
| Online Erişim: | https://ieeexplore.ieee.org/document/10703056/ |
| Etiketler: |
Etiket eklenmemiş, İlk siz ekleyin!
|
