Related Experiment Video
Updated: Jul 16, 2026

Assessing Human Spatial Navigation in a Virtual Space and its Sensitivity to Exercise
Published on: January 26, 2024
Guiding exploration by pre-existing knowledge without modifying reward
1Helsinki University of Technology, P.O. Box 5500, FIN-02015 HUT, Finland. Kary.Framling@hut.fi
Abstract:
Reinforcement learning is based on exploration of the environment and receiving reward that indicates which actions taken by the agent are good and which ones are bad. In many applications receiving even the first reward may require long exploration, during which the agent has no information about its progress. This paper presents an approach that makes it possible to use pre-existing knowledge about the task for guiding exploration through the state space. Concepts of short- and long-term memory combine guidance by pre-existing knowledge with reinforcement learning methods for value function estimation in order to make learning faster while allowing the agent to converge towards a good policy.
Related Concept Videos
Optimal Foraging
Cognitive Learning
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
Purposive Learning
Hindsight Biases
The Anchoring-and-Adjustment Heuristic
Incentive Theory: Pull Theory of Motivation
The theory differentiates between intrinsic and...

