Active Reward Learning for Co-Robotic Vision Based Exploration in Bandwidth Limited Environments

Jamieson, Stewart; How, Jonathan P.; Girdhar, Yogesh

doi:10.1109/ICRA40945.2020.9196922

Citation Details

Active Reward Learning for Co-Robotic Vision Based Exploration in Bandwidth Limited Environments

We present a novel POMDP problem formulation for a robot that must autonomously decide where to go to collect new and scientifically relevant images given a limited ability to communicate with its human operator. From this formulation, we derive constraints and design principles for the observation model, reward model, and communication strategy of such a robot, exploring techniques to deal with the very high-dimensional observation space and scarcity of relevant training data. We introduce a novel active reward learning strategy based on making queries to help the robot minimize path "regret" online, and evaluate it for suitability in autonomous visual exploration through simulations. We demonstrate that, in some bandwidth-limited environments, this novel regret-based criterion enables the robotic explorer to collect up to 17% more reward per mission than the next-best criterion. more »

Award ID(s):: 1734400

PAR ID:: 10206410

Author(s) / Creator(s):: Jamieson, Stewart; How, Jonathan P.; Girdhar, Yogesh

Date Published:: 2020-05-01

Journal Name:: 2020 IEEE International Conference on Robotics and Automation (ICRA)

Page Range / eLocation ID:: 1806 to 1812

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Conference Paper:
https://doi.org/10.1109/ICRA40945.2020.9196922

More Like this