Learning Adversarial Markov Decision Processes with Bandit Feedback and Unknown Transition
Name
jin20c.pdf
Description
Published version
Size
330.19 KB
Format
Adobe PDF
Checksum (MD5)
dfa0d651bd5036ada7ac709ee0266bad
Author(s) • • • •
Jin, Chi
Jin, Tiancheng
Luo, Haipeng
Sra, Suvrit
Yu, Tiancheng
Date Issued
2020
Journal
INTERNATIONAL CONFERENCE ON MACHINE LEARNING, VOL 119
Citation
Jin, Chi, Jin, Tiancheng, Luo, Haipeng, Sra, Suvrit and Yu, Tiancheng. 2020. "Learning Adversarial Markov Decision Processes with Bandit Feedback and Unknown Transition." INTERNATIONAL CONFERENCE ON MACHINE LEARNING, VOL 119, 119.
Version
Final published version
MIT Department
Massachusetts Institute of Technology. Institute for Data, Systems, and Society
Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science
Terms of Use
Article is made available in accordance with the publisher's policy and may be subject to US copyright law. Please refer to the publisher's site for terms of use.
Persistent DSpace Link
DOI of Published Version
https://proceedings.mlr.press/v119/jin20c.html