Christopher Thierauf, Matthias Scheutz. IEEE IROS 2024.
We can redesign the typical reinforcement learning pipeline to train faster and integrate with symbolic plans by training on the environment directly, not the robot within it.
You can read the paper here.
(more…)