Identifiability of Product of Experts Models

Gordon, Spencer L; Kant, Manav; Ma, Eric Y; Schulman, Leonard J; Staicu, Andrei C

Citation Details

Product of experts (PoE) are layered networks in which the value at each node is an AND (or product) of the values (possibly negated) at its inputs. These were introduced as a neural network architecture that can efficiently learn to generate high-dimensional data which satisfy many low-dimensional constraints---thereby allowing each individual expert to perform a simple task. PoEs have found a variety of applications in learning. We study the problem of identifiability of a product of experts model having a layer of binary latent variables, and a layer of binary observables that are iid conditional on the latents. The previous best upper bound on the number of observables needed to identify the model was exponential in the number of parameters. We show: (a) When the latents are uniformly distributed, the model is identifiable with a number of observables equal to the number of parameters (and hence best possible). (b) In the more general case of arbitrarily distributed latents, the model is identifiable for a number of observables that is still linear in the number of parameters (and within a factor of two of best-possible). The proofs rely on root interlacing phenomena for some special three-term recurrences. more »

Award ID(s):: 2321079

PAR ID:: 10508258

Author(s) / Creator(s):: Gordon, Spencer L; Kant, Manav; Ma, Eric Y; Schulman, Leonard J; Staicu, Andrei C

Publisher / Repository:: Proceedings of Machine Learning Research

Date Published:: 2024-05-02

Journal Name:: Proceedings of Machine Learning Research

Volume:: 238

ISSN:: 2640-3498

Format(s):: Medium: X

Location:: https://proceedings.mlr.press/v238/kant24a.html

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Conference Paper:
The DOI is not currently available.

More Like this