Learning poisson binomial distributions

Daskalakis, Constantinos; Diakonikolas, Ilias; Servedio, Rocco A.

The MIT Libraries is completing a major upgrade to DSpace@MIT. Starting May 5 2026, DSpace will remain functional, viewable, searchable, and downloadable, however, you will not be able to edit existing collections or add new material. We are aiming to have full functionality restored by May 18, 2026 but intermittent service interruptions may occur. Please email dspace-lib@mit.edu with any questions. Thank you for your patience as we implement this important upgrade.

Author(s)

Daskalakis, Constantinos; Diakonikolas, Ilias; Servedio, Rocco A.

DownloadDaskalakis-Learning Poisson.pdf (330.7Kb)

OPEN_ACCESS_POLICY

Terms of use

Creative Commons Attribution-Noncommercial-Share Alike 3.0 http://creativecommons.org/licenses/by-nc-sa/3.0/

Metadata

Show full item record

Abstract

We consider a basic problem in unsupervised learning: learning an unknown Poisson Binomial Distribution. A Poisson Binomial Distribution (PBD) over {0,1,...,n} is the distribution of a sum of n independent Bernoulli random variables which may have arbitrary, potentially non-equal, expectations. These distributions were first studied by S. Poisson in 1837 and are a natural n-parameter generalization of the familiar Binomial Distribution. Surprisingly, prior to our work this basic learning problem was poorly understood, and known results for it were far from optimal. We essentially settle the complexity of the learning problem for this basic class of distributions. As our main result we give a highly efficient algorithm which learns to ε-accuracy using O(1/ε[superscript 3]) samples independent of n. The running time of the algorithm is quasilinear in the size of its input data, i.e. ~O(log(n)/ε[superscript 3) bit-operations (observe that each draw from the distribution is a log(n)-bit string). This is nearly optimal since any algorithm must use Ω(1/ε[superscript 2) samples. We also give positive and negative results for some extensions of this learning problem.

Date issued

2012-05

URI

http://hdl.handle.net/1721.1/72345

Department

Massachusetts Institute of Technology. Computer Science and Artificial Intelligence Laboratory

Journal

Proceedings of the 44th symposium on Theory of Computing (STOC '12)

Publisher

Association for Computing Machinery (ACM)

Citation

Constantinos Daskalakis, Ilias Diakonikolas, and Rocco A. Servedio. 2012. Learning poisson binomial distributions. In Proceedings of the 44th symposium on Theory of Computing (STOC '12). ACM, New York, NY, USA, 709-728.

Version: Author's final manuscript

ISBN

978-1-4503-1245-5

Collections

MIT Open Access Articles