Journal of Neuroscience · 2007 · 406 citations · 26 references
NeuropsychologyReinforcement Learning SignalsAffective NeuroscienceReward-based Decision MakingCognitionAttentionSocial SciencesExperimental Decision MakingPublic HealthConditioningCognitive NeuroscienceCognitive ScienceBehavioral SciencesBehavioral NeuroscienceVisuomotor LearningReward SystemOperant BehaviorExperimental PsychologyDecision Making PerformanceExploration V ExploitationPredictive CodingAnticipatory ProcessNeuroeconomicsProcedural MemoryNeuroscienceDecision Neuroscience
Reinforcement learning models have advanced understanding of neural mechanisms underlying reward learning and decision‑making, yet human performance varies widely. The study aimed to identify the neural basis of individual differences in reward‑based decision‑making by comparing learners and nonlearners. Using a four‑armed bandit task and fMRI, 29 subjects were scanned while performing the task, and reinforcement learning models were applied to analyze neural activity. Learners exhibited robust prediction‑error signals in ventral and dorsal striatum, whereas nonlearners lacked such signals, and striatal prediction‑error magnitude correlated with behavioral performance, underscoring the role of dopaminergic striatal signaling in learning action selection.
The computational framework of reinforcement learning has been used to forward our understanding of the neural mechanisms underlying reward learning and decision-making behavior. It is known that humans vary widely in their performance in decision-making tasks. Here, we used a simple four-armed bandit task in which subjects are almost evenly split into two groups on the basis of their performance: those who do learn to favor choice of the optimal action and those who do not. Using models of reinforcement learning we sought to determine the neural basis of these intrinsic differences in performance by scanning both groups with functional magnetic resonance imaging. We scanned 29 subjects while they performed the reward-based decision-making task. Our results suggest that these two groups differ markedly in the degree to which reinforcement learning signals in the striatum are engaged during task performance. While the learners showed robust prediction error signals in both the ventral and dorsal striatum during learning, the nonlearner group showed a marked absence of such signals. Moreover, the magnitude of prediction error signals in a region of dorsal striatum correlated significantly with a measure of behavioral performance across all subjects. These findings support a crucial role of prediction error signals, likely originating from dopaminergic midbrain neurons, in enabling learning of action selection preferences on the basis of obtained rewards. Thus, spontaneously observed individual differences in decision making performance demonstrate the suggested dependence of this type of learning on the functional integrity of the dopaminergic striatal system in humans.
26
An Inventory for Measuring Depression
Aaron T. Beck · Archives of General Psychiatry · 1961 · 37.8K citations
Andrew G. Barto · IFAC Proceedings Volumes · 1998 · 3K citations