Q-based Variational Inverse Reinforcement Learning

By Ondrej Bajgar · Paper · cs.LG

The development of safe and beneficial AI requires that systems can learn and act in accordance with human preferences. However, explicitly specifying these preferences by hand is often infeasible. Inverse reinforcement learning (IRL) addresses this challenge by inferring prefere

Cs.lg

View original

HomeResourceLoading…