Learning poisson binomial distributions

Daskalakis, Constantinos; Diakonikolas, Ilias; Servedio, Rocco A.

dc.contributor.author	Daskalakis, Constantinos
dc.contributor.author	Diakonikolas, Ilias
dc.contributor.author	Servedio, Rocco A.
dc.date.accessioned	2021-11-05T11:05:38Z
dc.date.available	2021-11-05T11:05:38Z
dc.date.issued	2012
dc.identifier.uri	https://hdl.handle.net/1721.1/137414
dc.description.abstract	We consider a basic problem in unsupervised learning: learning an unknown \emph{Poisson Binomial Distribution}. A Poisson Binomial Distribution (PBD) over {0,1,…,n} is the distribution of a sum of n independent Bernoulli random variables which may have arbitrary, potentially non-equal, expectations. These distributions were first studied by S. Poisson in 1837 \cite{Poisson:37} and are a natural n-parameter generalization of the familiar Binomial Distribution. Surprisingly, prior to our work this basic learning problem was poorly understood, and known results for it were far from optimal. We essentially settle the complexity of the learning problem for this basic class of distributions. As our first main result we give a highly efficient algorithm which learns to $\eps$-accuracy (with respect to the total variation distance) using $\tilde{O}(1/\eps^3)$ samples \emph{independent of n}. The running time of the algorithm is \emph{quasilinear} in the size of its input data, i.e., $\tilde{O}(\log(n)/\eps^3)$ bit-operations. (Observe that each draw from the distribution is a log(n)-bit string.) Our second main result is a {\em proper} learning algorithm that learns to $\eps$-accuracy using $\tilde{O}(1/\eps^2)$ samples, and runs in time $(1/\eps)^{\poly (\log (1/\eps))} \cdot \log n$. This is nearly optimal, since any algorithm {for this problem} must use $\Omega(1/\eps^2)$ samples. We also give positive and negative results for some extensions of this learning problem to weighted sums of independent Bernoulli random variables.	en_US
dc.language.iso	en
dc.publisher	Association for Computing Machinery (ACM)	en_US
dc.relation.isversionof	10.1145/2213977.2214042	en_US
dc.rights	Creative Commons Attribution-Noncommercial-Share Alike	en_US
dc.rights.uri	http://creativecommons.org/licenses/by-nc-sa/4.0/	en_US
dc.source	arXiv	en_US
dc.title	Learning poisson binomial distributions	en_US
dc.type	Article	en_US
dc.identifier.citation	Daskalakis, Constantinos, Diakonikolas, Ilias and Servedio, Rocco A. 2012. "Learning poisson binomial distributions."
dc.contributor.department	Massachusetts Institute of Technology. Computer Science and Artificial Intelligence Laboratory	en_US
dc.eprint.version	Author's final manuscript	en_US
dc.type.uri	http://purl.org/eprint/type/ConferencePaper	en_US
eprint.status	http://purl.org/eprint/status/NonPeerReviewed	en_US
dc.date.updated	2019-05-15T17:55:03Z
dspace.date.submission	2019-05-15T17:55:03Z
mit.metadata.status	Authority Work and Publication Information Needed	en_US

Files in this item

Name:: 1107.2702.pdf
Size:: 384.9Kb
Format:: PDF
Description:: Accepted version

View/Open

This item appears in the following Collection(s)

MIT Open Access Articles

Show simple item record