Bootstrapping EM via EM and Convergence Analysis in the Naive Bayes Model

Daskalakis, Constantinos; Tzamos, Christos; Zampetakis, Manolis

Citation Details

We study the convergence properties of the Expectation-Maximization algorithm in the Naive Bayes model. We show that EM can get stuck in regions of slow convergence, even when the features are binary and i.i.d. conditioning on the class label, and even under random (i.e. non worst-case) initialization. In turn, we show that EM can be bootstrapped in a pre-training step that computes a good initialization. From this initialization we show theoretically and experimentally that EM converges exponentially fast to the true model parameters. Our bootstrapping method amounts to running the EM algorithm on appropriately centered iterates of small magnitude, which as we show corresponds to effectively performing power iteration on the covariance matrix of the mixture model, although power iteration is performed under the hood by EM itself. As such, we call our bootstrapping approach “power EM.” Specifically for the case of two binary features, we show global exponentially fast convergence of EM, even without bootstrapping. Finally, as the Naive Bayes model is quite expressive, we show as corollaries of our convergence results that the EM algorithm globally converges to the true model parameters for mixtures of two Gaussians, recovering recent results of Xu et al.’2016 and Daskalakis et al. 2017. more »

Award ID(s):: 1741137

PAR ID:: 10079722

Author(s) / Creator(s):: Daskalakis, Constantinos; Tzamos, Christos; Zampetakis, Manolis

Date Published:: 2018-04-16

Journal Name:: Proceedings of the International Workshop on Artificial Intelligence and Statistics

ISSN:: 1525-531X

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Conference Paper:
The DOI is not currently available.

More Like this