Tight Bounds on the Hardness of Learning Simple Nonparametric Mixtures

Tai, Wai_Ming; Aragam, Bryon

Citation Details

We study the problem of learning nonparametric distributions in a finite mixture, and establish tight bounds on the sample complexity for learning the component distributions in such models.Namely, we are given i.i.d. samples from a pdf f where f=w1f1+w2f2,w1+w2=1,w1,w2>0 and we are interested in learning each component fi .Without any assumptions on fi , this problem is ill-posed.In order to identify the components fi , we assume that each fi can be written as a convolution of a Gaussian and a compactly supported density νi with supp(ν1)∩supp(ν2)=∅ .Our main result shows that (1ε)Ω(loglog1ε) samples are required for estimating each fi . The proof relies on a quantitative Tauberian theorem that yields a fast rate of approximation with Gaussians, which may be of independent interest. To show this is tight, we also propose an algorithm that uses (1ε)O(loglog1ε) samples to estimate each fi . Unlike existing approaches to learning latent variable models based on moment-matching and tensor methods, our proof instead involves a delicate analysis of an ill-conditioned linear system via orthogonal functions.Combining these bounds, we conclude that the optimal sample complexity of this problem properly lies in between polynomial and exponential, which is not common in learning theory. more »

Award ID(s):: 1956330

PAR ID:: 10542254

Author(s) / Creator(s):: Tai, Wai_Ming; Aragam, Bryon

Publisher / Repository:: Proceedings of Thirty Sixth Conference on Learning Theory

Date Published:: 2023-07-15

Volume:: 195

Page Range / eLocation ID:: 2849-2849

Subject(s) / Keyword(s):: learning theory nonparametric statistics sample complexity lower bounds deconvolution

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Conference Paper:
The DOI is not currently available.

More Like this