When Does On-Policy Interaction Help? Representational Tradeoffs in Value-Based Imitation Learning
By Luca Viano · Paper · cs.LG
Imitation learning (IL)---training an agent to replicate expert behavior from demonstrations---underpins applications from robotics to language model training. Standard approaches such as Behavior Cloning (BC) are known to suffer from compounding errors and performance plateaus,