Dropout Training is Distributionally Robust Optimal

Blanchet, Jose; Montiel-Olea, José Luis; Nguyen, Viet Anh; Zhang, Xuhui

Citation Details

This paper shows that dropout training in generalized linear models is the minimax solution of a two-player, zero-sum game where an adversarial nature corrupts a statistician's covariates using a multiplicative nonparametric errors-in-variables model. In this game, nature's least favorable distribution is dropout noise, where nature independently deletes entries of the covariate vector with some fixed probability δ. This result implies that dropout training indeed provides out-of-sample expected loss guarantees for distributions that arise from multiplicative perturbations of in-sample data. The paper makes a concrete recommendation on how to select the tuning parameter δ. The paper also provides a novel, parallelizable, unbiased multi-level Monte Carlo algorithm to speed-up the implementation of dropout training. Our algorithm has a much smaller computational cost compared to the naive implementation of dropout, provided the number of data points is much smaller than the dimension of the covariate vector. more »

Award ID(s):: 1915967

PAR ID:: 10483201

Author(s) / Creator(s):: Blanchet, Jose; Montiel-Olea, José Luis; Nguyen, Viet Anh; Zhang, Xuhui

Publisher / Repository:: Journal of Machine Learning Research

Date Published:: 2023-06-01

Journal Name:: Journal of machine learning research

Volume:: 24

Issue:: 180

ISSN:: 1533-7928

Page Range / eLocation ID:: 1-60

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Journal Article:
The DOI is not currently available.

More Like this