Do learning rates adapt to the distribution of rewards?
Name
13423_2014_790_ReferencePDF.pdf
Size
97.16 KB
Format
Adobe PDF
Checksum (MD5)
e1e9f827f6a249bc340bb1a3e092db4b
Author(s)
Gershman, Samuel J.
Date Issued
January 2015
Journal
Psychonomic Bulletin & Review
Publisher
Springer US
Citation
Gershman, Samuel J. “Do Learning Rates Adapt to the Distribution of Rewards?” Psychonomic Bulletin & Review 22.5 (2015): 1320–1327.
Version
Author's final manuscript
Abstract
Studies of reinforcement learning have shown that humans learn differently in response to positive and negative reward prediction errors, a phenomenon that can be captured computationally by positing asymmetric learning rates. This asymmetry, motivated by neurobiological and cognitive considerations, has been invoked to explain learning differences across the lifespan as well as a range of psychiatric disorders. Recent theoretical work, motivated by normative considerations, has hypothesized that the learning rate asymmetry should be modulated by the distribution of rewards across the available options. In particular, the learning rate for negative prediction errors should be higher than the learning rate for positive prediction errors when the average reward rate is high, and this relationship should reverse when the reward rate is low. We tested this hypothesis in a series of experiments. Contrary to the theoretical predictions, we found that the asymmetry was largely insensitive to the average reward rate; instead, the dominant pattern was a higher learning rate for negative than for positive prediction errors, possibly reflecting risk aversion.
MIT Department
Massachusetts Institute of Technology. Department of Brain and Cognitive Sciences
Terms of Use
Article is made available in accordance with the publisher's policy and may be subject to US copyright law. Please refer to the publisher's site for terms of use.
Persistent DSpace Link
DOI of Published Version
https://doi.org/10.3758/s13423-014-0790-3