Safe Robot Learning in Assistive Devices through Neural Network Repair

Keyvan Majd, Geoffrey Mitchell

Citation Details

Assistive robotic devices are a particularly promising field of application for neural networks (NN) due to the need for personalization and hard-to-model human-machine interaction dynamics. However, NN based estimators and controllers may produce potentially unsafe outputs over previously unseen data points. In this paper, we introduce an algorithm for updating NN control policies to satisfy a given set of formal safety constraints, while also optimizing the original loss function. Given a set of mixed-integer linear constraints, we define the NN repair problem as a Mixed Integer Quadratic Program (MIQP). In extensive experiments, we demonstrate the efficacy of our repair method in generating safe policies for a lower-leg prosthesis. more »

Award ID(s):: 1932189

PAR ID:: 10474224

Author(s) / Creator(s):: Keyvan Majd, Geoffrey Mitchell

Publisher / Repository:: Proc. of Machine Learning Resaerch

Date Published:: 2023-08-23

Journal Name:: Proceedings of Machine Learning Research

Volume:: 205

ISSN:: 2640-3498

Page Range / eLocation ID:: 2148-2158

Format(s):: Medium: X

Location:: https://proceedings.mlr.press/v205/majd23a.html

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Conference Paper:
The DOI is not currently available.

More Like this