Strong mixed-integer programming formulations for trained neural networks
Name
10107_2020_1474_ReferencePDF.pdf
Size
567.03 KB
Format
Adobe PDF
Checksum (MD5)
637bf16287aafc7a2b24b180f37b30cf
Author(s) • • • •
Anderson, Ross
Huchette, Joey
Ma, Will
Tjandraatmadja, Christian
Vielma, Juan P
Date Issued
February 13, 2020
Publisher
Springer Berlin Heidelberg
Version
Author's final manuscript
Abstract
Abstract
We present strong mixed-integer programming (MIP) formulations for high-dimensional piecewise linear functions that correspond to trained neural networks. These formulations can be used for a number of important tasks, such as verifying that an image classification network is robust to adversarial inputs, or solving decision problems where the objective function is a machine learning model. We present a generic framework, which may be of independent interest, that provides a way to construct sharp or ideal formulations for the maximum of d affine functions over arbitrary polyhedral input domains. We apply this result to derive MIP formulations for a number of the most popular nonlinear operations (e.g. ReLU and max pooling) that are strictly stronger than other approaches from the literature. We corroborate this computationally, showing that our formulations are able to offer substantial improvements in solve time on verification tasks for image classification networks.
MIT Department
Sloan School of Management
Terms of Use
Article is made available in accordance with the publisher's policy and may be subject to US copyright law. Please refer to the publisher's site for terms of use.
Persistent DSpace Link
DOI of Published Version
https://doi.org/10.1007/s10107-020-01474-5