Related Experiment Video
Updated: Jun 16, 2026

Measuring Delay Discounting in Humans Using an Adjusting Amount Task
Published on: January 9, 2016
Hyperbolically discounted temporal difference learning
William H Alexander1, Joshua W Brown
1Department of Psychological and Brain Sciences, Indiana University, Bloomington, IN 47405, USA. wialexan@indiana.edu
Abstract:
Hyperbolic discounting of future outcomes is widely observed to underlie choice behavior in animals. Additionally, recent studies (Kobayashi & Schultz, 2008) have reported that hyperbolic discounting is observed even in neural systems underlying choice. However, the most prevalent models of temporal discounting, such as temporal difference learning, assume that future outcomes are discounted exponentially. Exponential discounting has been preferred largely because it can be expressed recursively, whereas hyperbolic discounting has heretofore been thought not to have a recursive definition. In this letter, we define a learning algorithm, hyperbolically discounted temporal difference (HDTD) learning, which constitutes a recursive formulation of the hyperbolic model.
Related Concept Videos
Linear Approximation in Time Domain
For a simple pendulum with a mass evenly distributed along its length and the center of mass located at half the pendulum's length, the...
Difference from Background: Limit of Detection
The LOD indicates the presence or absence...
Hindsight Biases
Observational Learning

