Related Experiment Video
Updated: Feb 16, 2026

A Prediction Error-driven Retrieval Procedure for Destabilizing and Rewriting Maladaptive Reward Memories in Hazardous Drinkers
Published on: January 5, 2018
Dual reward prediction components yield Pavlovian sign- and goal-tracking
Sivaramakrishnan Kaveri1, Hiroyuki Nakahara2
1Lab for Integrated Theoretical Neuroscience, RIKEN BSI, Wako, Japan; Dept. of Computational Intelligence and Systems Science, Tokyo Institute of Technology, Yokohama, Japan.
Abstract:
Reinforcement learning (RL) has become a dominant paradigm for understanding animal behaviors and neural correlates of decision-making, in part because of its ability to explain Pavlovian conditioned behaviors and the role of midbrain dopamine activity as reward prediction error (RPE). However, recent experimental findings indicate that dopamine activity, contrary to the RL hypothesis, may not signal RPE and differs based on the type of Pavlovian response (e.g. sign- and goal-tracking responses). In this study, we address this discrepancy by introducing a new neural correlate for learning reward predictions; the correlate is called "cue-evoked reward". It refers to a recall of reward evoked by the cue that is learned through simple cue-reward associations. We introduce a temporal difference learning model, in which neural correlates of the cue itself and cue-evoked reward underlie learning of reward predictions. The animal's reward prediction supported by these two correlates is divided into sign and goal components respectively. We relate the sign and goal components to approach responses towards the cue (i.e. sign-tracking) and the food-tray (i.e. goal-tracking) respectively. We found a number of correspondences between simulated models and the experimental findings (i.e. behavior and neural responses). First, the development of modeled responses is consistent with those observed in the experimental task. Second, the model's RPEs were similar to dopamine activity in respective response groups. Finally, goal-tracking, but not sign-tracking, responses rapidly emerged when RPE was restored in the simulated models, similar to experiments with recovery from dopamine-antagonist. These results suggest two complementary neural correlates, corresponding to the cue and its evoked reward, form the basis for learning reward predictions in the sign- and goal-tracking rats.
Related Concept Videos
Psychosis: Goals of Pharmacotherapy
ATP Yield
The ETC is embedded in the inner mitochondrial membrane and is comprised of four main protein complexes and an ATP synthase. NADH and FADH2 pass electrons to these complexes, which pump protons into the intermembrane space. This distribution of...
Reaction Yield
Sign Convention
The normal force acts perpendicular to the beam's cross-section and can...
Signs of Puberty
Introduction to the Sign Test

