$R^3$ trains robots to use natural-language reasoning during manipulation
An arXiv preprint introduces $R^3$, a post-training method that uses free-form language reasoning to guide robotic manipulation policies. The authors report gains on two controlled benchmarks, while leaving real-world performance and the size of those gains unspecified.