Limits of Private Learning with Access to Public Data

Alon, N.; Bassily, R.; and Moran, S.

Citation Details

We consider learning problems where the training set consists of two types of examples: private and public. The goal is to design a learning algorithm that satisfies differential privacy only with respect to the private examples. This setting interpolates between private learning (where private) and classical learning (where all examples are public). We study the limits of learning in this setting in terms of private and public sample complexities. We show that any hypothesis class of VC-dimension d can be agnostically learned up to an excess error of α using only (roughly) d/α public examples and d/α2 private labeled examples. This result holds even when the public examples are unlabeled. This gives a quadratic improvement over the standard d/α2 upper bound on the public sample complexity (where private examples can be ignored altogether if the public examples are labeled). Furthermore, we give a nearly matching lower bound, which we prove via a generic reduction from this setting to the one of private learning without public data. more »

Award ID(s):: 1855464

PAR ID:: 10144213

Author(s) / Creator(s):: Alon, N.; Bassily, R.; and Moran, S.

Date Published:: 2019-10-25

Journal Name:: Advances in neural information processing systems

Issue:: 2019

ISSN:: 1049-5258

Page Range / eLocation ID:: 10342-10352

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Journal Article:
The DOI is not currently available.

More Like this