Show simple item record

dc.contributor.authorBertsekas, Dimitri P
dc.date.accessioned2018-06-12T19:20:01Z
dc.date.available2018-06-12T19:20:01Z
dc.date.issued2017-01
dc.date.submitted2017-05
dc.identifier.issn1052-6234
dc.identifier.issn1095-7189
dc.identifier.urihttp://hdl.handle.net/1721.1/116281
dc.description.abstractWe consider challenging dynamic programming models where the associated Bellman equation, and the value and policy iteration algorithms commonly exhibit complex and even pathological behavior. Our analysis is based on the new notion of regular policies. These are policies that are well-behaved with respect to value and policy iteration, and are patterned after proper policies, which are central in the theory of stochastic shortest path problems. We show that the optimal cost function over regular policies may have favorable value and policy iteration properties, which the optimal cost function over all policies need not have. We accordingly develop a unifying methodology to address long standing analytical and algorithmic issues in broad classes of undiscounted models, including stochastic and minimax shortest path problems, as well as positive cost, negative cost, risk-sensitive, and multiplicative cost problems.en_US
dc.publisherSociety for Industrial & Applied Mathematics (SIAM)en_US
dc.relation.isversionofhttp://dx.doi.org/10.1137/16M1090946en_US
dc.rightsArticle is made available in accordance with the publisher's policy and may be subject to US copyright law. Please refer to the publisher's site for terms of use.en_US
dc.sourceSociety for Industrial and Applied Mathematicsen_US
dc.titleRegular Policies in Abstract Dynamic Programmingen_US
dc.typeArticleen_US
dc.identifier.citationBertsekas, Dimitri P. “Regular Policies in Abstract Dynamic Programming.” SIAM Journal on Optimization 27, 3 (January 2017): 1694–1727 © 2017 Society for Industrial and Applied Mathematicsen_US
dc.contributor.departmentMassachusetts Institute of Technology. Department of Electrical Engineering and Computer Scienceen_US
dc.contributor.mitauthorBertsekas, Dimitri P
dc.relation.journalSIAM Journal on Optimizationen_US
dc.eprint.versionFinal published versionen_US
dc.type.urihttp://purl.org/eprint/type/JournalArticleen_US
eprint.statushttp://purl.org/eprint/status/PeerRevieweden_US
dc.date.updated2018-05-11T14:03:15Z
dspace.orderedauthorsBertsekas, Dimitri P.en_US
dspace.embargo.termsNen_US
dc.identifier.orcidhttps://orcid.org/0000-0001-6909-7208
mit.licensePUBLISHER_POLICYen_US


Files in this item

Thumbnail

This item appears in the following Collection(s)

Show simple item record