Time representation in reinforcement learning models of the basal ganglia

Samuel J Gershman1, Ahmed A Moustafa2, Elliot A Ludvig3

  • 1Department of Brain and Cognitive Sciences, Massachusetts Institute of Technology Cambridge, MA, USA.

Related Concept Videos

State Space Representation01:27

State Space Representation

The frequency-domain technique, commonly used in analyzing and designing feedback control systems, is effective for linear, time-invariant systems. However, it falls short when dealing with nonlinear, time-varying, and multiple-input multiple-output systems. The time-domain or state-space approach addresses these limitations by utilizing state variables to construct simultaneous, first-order differential equations, known as state equations, for an nth-order system.
Consider an RLC circuit, a...
785
Reinforcement Schedules01:24

Reinforcement Schedules

Positive reinforcement is a powerful method for teaching new behaviors to both animals and humans. B.F. Skinner demonstrated this with his experiments using rats in a Skinner box. When a rat pressed a lever, it received a food pellet. This immediate reward encouraged the rat to repeat the behavior. This method, where a reward follows every instance of the behavior, is known as continuous reinforcement. It is highly effective for establishing new behaviors quickly.
Once a behavior is learned,...
740
Long-term Potentiation01:35

Long-term Potentiation

Long-term potentiation, or LTP, is one of the ways by which synaptic plasticity—changes in the strength of chemical synapses—can occur in the brain. LTP is the process of synaptic strengthening that occurs over time between pre- and postsynaptic neuronal connections. The synaptic strengthening of LTP works in opposition to the synaptic weakening of long-term depression (LTD) and together are the main mechanisms that underlie learning and memory.
51.6K
Timing and Consequences on Behavior01:08

Timing and Consequences on Behavior

In operant conditioning, the timing of reinforcement is crucial. For animals like rats and cats, immediate reinforcement (within a few seconds) is much more effective than delayed reinforcement. For example, a food reward for a rat needs to follow within 30 seconds of pressing a bar to be effective. 
Humans, however, can respond to delayed reinforcers. We often make decisions between immediate small rewards and delayed larger rewards. This ability to delay gratification is a significant...
927
Diencephalon: Thalamus and Information Relay01:27

Diencephalon: Thalamus and Information Relay

The thalamus, often called “the gateway to the cerebral cortex,” is vital in processing and directing sensory and motor signals throughout the brain. Almost all inputs destined for the cerebral cortex, except for olfactory signals, are relayed through the thalamus. The thalamus is  a sophisticated relay station, channeling information from various brain regions to the cerebral cortex, as well as a filter, prioritizing certain signals over others based on current physiological...
4.9K