dc.contributor.author | Liang, Tengyuan | |
dc.contributor.author | Rakhlin, Alexander | |
dc.date.accessioned | 2021-12-03T15:51:24Z | |
dc.date.available | 2021-12-03T15:51:24Z | |
dc.date.issued | 2020 | |
dc.identifier.uri | https://hdl.handle.net/1721.1/138308 | |
dc.description.abstract | © Institute of Mathematical Statistics, 2020. In the absence of explicit regularization, Kernel “Ridgeless” Regression with nonlinear kernels has the potential to fit the training data perfectly. It has been observed empirically, however, that such interpolated solutions can still generalize well on test data. We isolate a phenomenon of implicit regularization for minimum-norm interpolated solutions which is due to a combination of high dimensionality of the input data, curvature of the kernel function and favorable geometric properties of the data such as an eigenvalue decay of the empirical covariance and kernel matrices. In addition to deriving a data-dependent upper bound on the out-of-sample error, we present experimental evidence suggesting that the phenomenon occurs in the MNIST dataset. | en_US |
dc.language.iso | en | |
dc.publisher | Institute of Mathematical Statistics | en_US |
dc.relation.isversionof | 10.1214/19-AOS1849 | en_US |
dc.rights | Creative Commons Attribution-Noncommercial-Share Alike | en_US |
dc.rights.uri | http://creativecommons.org/licenses/by-nc-sa/4.0/ | en_US |
dc.source | arXiv | en_US |
dc.title | Just interpolate: Kernel “Ridgeless” regression can generalize | en_US |
dc.type | Article | en_US |
dc.identifier.citation | Liang, Tengyuan and Rakhlin, Alexander. 2020. "Just interpolate: Kernel “Ridgeless” regression can generalize." Annals of Statistics, 48 (3). | |
dc.contributor.department | Massachusetts Institute of Technology. Institute for Data, Systems, and Society | |
dc.contributor.department | Statistics and Data Science Center (Massachusetts Institute of Technology) | |
dc.contributor.department | Massachusetts Institute of Technology. Department of Brain and Cognitive Sciences | |
dc.contributor.department | Massachusetts Institute of Technology. Laboratory for Information and Decision Systems | |
dc.relation.journal | Annals of Statistics | en_US |
dc.eprint.version | Author's final manuscript | en_US |
dc.type.uri | http://purl.org/eprint/type/JournalArticle | en_US |
eprint.status | http://purl.org/eprint/status/PeerReviewed | en_US |
dc.date.updated | 2021-12-03T15:41:20Z | |
dspace.orderedauthors | Liang, T; Rakhlin, A | en_US |
dspace.date.submission | 2021-12-03T15:41:22Z | |
mit.journal.volume | 48 | en_US |
mit.journal.issue | 3 | en_US |
mit.license | OPEN_ACCESS_POLICY | |
mit.metadata.status | Authority Work and Publication Information Needed | en_US |