☆ 4.3 Article

Learning the opportunity cost of time in a patch-foraging task

COGNITIVE AFFECTIVE & BEHAVIORAL NEUROSCIENCE (2015)

期刊

COGNITIVE AFFECTIVE & BEHAVIORAL NEUROSCIENCE

卷 15, 期 4, 页码 837-853

出版社

SPRINGER

DOI: 10.3758/s13415-015-0350-y

关键词

Computational model; Decision making; Dopamine; Reward; Patchforaging; Reinforcement learning

类别

Behavioral Sciences Neurosciences

资金

Human Frontiers Science Program [RGP0036/2009-C]
National Institute of Mental Health [R01MH087882]
McDonnell Foundation

向作者/读者索取更多资源

Protocol

社区支持

Reagent

社区支持

摘要

Although most decision research concerns choice between simultaneously presented options, in many situations options are encountered serially, and the decision is whether to exploit an option or search for a better one. Such problems have a rich history in animal foraging, but we know little about the psychological processes involved. In particular, it is unknown whether learning in these problems is supported by the well-studied neurocomputational mechanisms involved in more conventional tasks. We investigated how humans learn in a foraging task, which requires deciding whether to harvest a depleting resource or switch to a replenished one. The optimal choice (given by the marginal value theorem; MVT) requires comparing the immediate return from harvesting to the opportunity cost of time, which is given by the long-run average reward. In two experiments, we varied opportunity cost across blocks, and subjects adjusted their behavior to blockwise changes in environmental characteristics. We examined how subjects learned their choice strategies by comparing choice adjustments to a learning rule suggested by the MVT (in which the opportunity cost threshold is estimated as an average over previous rewards) and to the predominant incremental-learning theory in neuroscience, temporal-difference learning (TD). Trial-by-trial decisions were explained better by the MVT threshold-learning rule. These findings expand on the foraging literature, which has focused on steady-state behavior, by elucidating a computational mechanism for learning in switching tasks that is distinct from those used in traditional tasks, and suggest connections to research on average reward rates in other domains of neuroscience.

Learning the opportunity cost of time in a patch-foraging task

期刊

COGNITIVE AFFECTIVE & BEHAVIORAL NEUROSCIENCE

出版社

SPRINGER

关键词

类别

资金

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

Learning the opportunity cost of time in a patch-foraging task

期刊

COGNITIVE AFFECTIVE & BEHAVIORAL NEUROSCIENCE

出版社

SPRINGER

关键词

类别

资金

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

导出引文

分享论文