Updated · 1 episodes · 1 show · 1 source notes
Reward Prediction Error Learning
Definition
Reward prediction error learning is the episode’s frame for how the brain updates future motivation when an outcome is better, worse, or equal to what was expected.
Current Synthesis
In Leverage Dopamine to Overcome Procrastination & Optimize Effort, reward prediction error connects dopamine to learning rather than only to pleasure. Dopamine can rise in anticipation of a desired outcome, then update based on the difference between expected and obtained reward. Cues that reliably appear between wanting and reward can become part of the learned pursuit sequence, so environments, reminders, timing, and intermediate signals can start to pull behavior before the final reward arrives. This makes motivation partly historical: future effort depends on what previous pursuit taught the system to expect.
Key Claims
- Dopamine signals can occur before reward when an organism anticipates a desired outcome.
- Better-than-expected outcomes can increase future pursuit, while worse-than-expected outcomes can reduce it.
- Cues between wanting and receiving reward can become learned signals that shape behavior.
- Reward prediction error makes motivation sensitive to expectation, surprise, and prior outcomes.
- Procrastination, craving, and pursuit can be influenced by cue learning as well as conscious goals.
Evidence
- Anticipation signal: Leverage Dopamine to Overcome Procrastination & Optimize Effort says dopamine is released in anticipation, not only at reward receipt.
- Expectation update: Leverage Dopamine to Overcome Procrastination & Optimize Effort describes reward prediction error as the comparison between experienced and expected reward.
- Cue learning: Leverage Dopamine to Overcome Procrastination & Optimize Effort says cues between wanting and reward can enter reward-contingent learning.
- Motivation consequence: Leverage Dopamine to Overcome Procrastination & Optimize Effort uses the mechanism to explain future pursuit, craving, and effort allocation.
Counterevidence & Qualifications
This page captures the episode’s applied explanation, not a full computational reinforcement-learning model. It should not imply that all learning, addiction, habit formation, or procrastination is reducible to one dopamine signal or one prediction-error equation.
What Changed
- Created a source-grounded page for dopamine-linked anticipation, expectation, cue learning, and future motivation.
Related Concepts
- Dopamine Peak-Trough Baseline - baseline and trough context in which reward updates affect motivation.
- Dopamine Wanting Loop / 多巴胺渴爱循环 - wanting and pursuit branch shaped by learned cues.
- Motivation Reward-Effort Calculation - action-selection frame that uses expected reward and prior outcomes.
- Attention Capacity Selection - attention branch because salient cues compete for control.
- Effort As Reward - motivation-design branch where effort can become part of the learned reward signal.