QR koda

Gamma-Regression-Based Inverse Reinforcement Learning From Suboptimal Demonstrations

Inverse reinforcement learning (IRL) is a technique that estimates the intention of an expert who acts optimally on a specific intention, as a reward from demonstration (i.e., recorded data of the expert’s behavior). Traditional IRL algorithms are based on the assumption that expert demonstra...

Popoln opis

Shranjeno v:
Bibliografske podrobnosti
Principais autores: Daiko Kishikawa, Sachiyo Arai
Format: Artigo
Jezik:Inglês
Izdano: IEEE 2024-01-01
Serija:IEEE Access
Teme:
Online dostop:https://ieeexplore.ieee.org/document/10669006/
Oznake: Označite
Brez oznak, prvi označite!