$R^3$: Training Robots to Reason in Natural Language via Reinforcement Learning
By Lehong Wu · Paper · cs.RO
Reasoning in language allows foundation models to spend more test-time compute on hard problems, such as those requiring decomposition, constraint tracking, and prediction of future consequences. Whether this mechanism can improve robotic manipulation remains unclear, where long-